arXiv:2609.05553v1 Announce Type: new Abstract: Agent memory allows LLM agents to use earlier interactions when answering new queries. Existing methods…
Category: cs.AI updates on arXiv.org
Deep belief networks are exact
arXiv:2609.05572v1 Announce Type: new Abstract: We prove that every strictly positive probability distribution on \(\{-1,1\}^n\) is represented exactly by…
Planning and Scheduling Business Processes under Control-Flow Uncertainty
arXiv:2609.05578v1 Announce Type: new Abstract: Scheduling activities in business processes can improve efficiency (e.g., reduce makespan), but is…
EnvCraft: Synthesizing Executable Environments in Agentic RL for Claw-like Agent
arXiv:2609.05576v1 Announce Type: new Abstract: The paradigm of LLMs has rapidly shifted from passive language interfaces to autonomous Claw-like agents…
Reasoning-Aware Compression: Identifying and Protecting Vulnerable Reasoning Circuits for Energy-Efficient LLM Deployment
arXiv:2609.05512v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) impose substantial energy costs during deployment, yet current compression…
The Failure Happens Before the Drift: The Social Dynamics of Values in LLM Agent Societies
arXiv:2609.05514v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents are increasingly used as proxies for human participants in social…
When and What to Teach: Budget-Aware Online Adaptation for Web Agents
arXiv:2609.05513v1 Announce Type: new Abstract: Web agents have achieved significant success in automating complex internet tasks but deploying them in…
Beyond “AI Helps Humans”: Decision-Targeted Evaluation Design for Human-Agent Teams in the Agentic Era
arXiv:2609.05527v1 Announce Type: new Abstract: Wherever a coding agent works under engineer supervision, or a clinical model assists a radiologist, the…
SCAFFOLD: Self-Improving Web Agents via Recursive Parametric Skill Abstraction
arXiv:2609.05511v1 Announce Type: new Abstract: Web agents need to navigate visually rich, long-horizon interfaces that change across sites, yet most…
SciLitBench: Benchmark and Design Principles for LLM-Powered Systematic Literature Reviews
arXiv:2609.05505v1 Announce Type: new Abstract: Systematic reviews require sustained human judgment across thousands of records, yet existing evaluations…
