arXiv:2609.03787v1 Announce Type: new Abstract: AI agents increasingly gather evidence, invoke tools, apply constraints, and produce decisions that people…
Tag: cs.AI updates on arXiv.org
SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation
arXiv:2609.03806v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) generation is attracting increasing attention as generative models improve…
Rethinking World Models for Safety-Critical Embodied Systems
arXiv:2609.03774v1 Announce Type: new Abstract: World models have progressed from compact latent dynamics to generative, controllable, and interactive…
Govern the Model, Not Only the Data: Storage, Circulation, and Learning in Creative AI
arXiv:2609.03800v1 Announce Type: new Abstract: Federated learning is increasingly presented as a privacy-preserving advance: personal data remain on the…
Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation
arXiv:2609.03727v1 Announce Type: new Abstract: Large language model agents can plan, invoke tools, and modify external states, yet most systems still…
Counterfactual Routing Using Integer Programming with Constraint Generation
arXiv:2609.03707v1 Announce Type: new Abstract: We present our submission to the IJCAI 2025 ‘Counterfactual Routing Competition’ (CRC 25). The goal of the…
Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study
arXiv:2609.03702v1 Announce Type: new Abstract: General-purpose code embeddings power tools for code search, classification, and retrieval. Compact…
Artificial Intelligence for Energy Optimization in Data Centers
arXiv:2609.03716v1 Announce Type: new Abstract: Data centers are increasingly optimized by artificial intelligence and, at the same time, increasingly…
SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation
arXiv:2609.03753v1 Announce Type: new Abstract: As large language models (LLMs) become increasingly capable, the long-term value of AI systems depends not…
HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews
arXiv:2609.03580v1 Announce Type: new Abstract: The growing scale of academic peer review has motivated the use of Large Language Models (LLMs) as review…
