arXiv:2608.22610v1 Announce Type: new Abstract: Agent skills, structured artifacts distilled from interaction trajectories and dynamically reused from…
Tag: cs.AI updates on arXiv.org
CAI-DLLM: Convergence Aware Inference for Diffusion Language Models
arXiv:2608.22646v1 Announce Type: new Abstract: Diffusion language models can generate many tokens in parallel, but they still require repeated denoising…
A-CPES: A Reference Framework for Agentic AI in Cyber-Physical Energy Systems
arXiv:2608.22672v1 Announce Type: new Abstract: Energy system operation contains a loop of work that automation has never taken over: posing the…
CausalCache: Conditional High-Fidelity Restoration for Long-Horizon GUI Agents
arXiv:2608.22577v1 Announce Type: new Abstract: Long-horizon GUI agents can retain a complete interaction trace cheaply as textual action records, but…
CONTRAMEM: Learning Self-Evolving Procedural Memory from Contrasting Multi-Model Trajectories
arXiv:2608.22533v1 Announce Type: new Abstract: Autonomous computer-use agents are increasingly applied to long-horizon tasks requiring coordinated…
STAGE: Stateful Translation to Agentic Graph Execution with Policy-Scoped Context and Deterministic Control
arXiv:2608.22538v1 Announce Type: new Abstract: Policy-governed agents must interpret case evidence while following an authorized procedure. We present…
ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation
arXiv:2608.22559v1 Announce Type: new Abstract: Rubrics aim to make language-model evaluation transparent by decomposing response quality into…
Scaling Curriculum Learning For Autonomous Driving
arXiv:2608.22549v1 Announce Type: new Abstract: Batched simulators for autonomous driving have recently enabled training reinforcement learning (RL)…
Small Reasoning Models are Instruction Followers in Function Calling
arXiv:2608.22472v1 Announce Type: new Abstract: Function calling represents the core capability of agentic large language models (LLMs). Existing research…
HANSARD: A Reference Architecture for Forensic Readiness, Runtime Witnessing, and Graded Attribution in Autonomous Multi-Agent AI Systems
arXiv:2608.22512v1 Announce Type: new Abstract: Autonomous multi-agent systems nowadays act in finance, software supply chains, and security operations.…
