arXiv:2608.08471v1 Announce Type: new Abstract: Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new…
Category: cs.AI updates on arXiv.org
Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses
arXiv:2608.08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the…
TRACE-Memory: Public-Conditioned Retrieval and Utility-Aware Evidence Admission for Personalized Generation
arXiv:2608.08446v1 Announce Type: new Abstract: Personalized generation systems retrieve user history by request–memory relevance and inject it into the…
Estimating Uncertainty in Galaxy Morphology Classification
arXiv:2608.08398v1 Announce Type: new Abstract: Astronomers classify galaxy morphology to investigate cosmic evolution. While deep foundation models are…
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
arXiv:2608.08389v1 Announce Type: new Abstract: Long-horizon research agents solve open-ended tasks through iterative retrieval, aggregation, and…
Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR Perspective
arXiv:2608.08445v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely regarded as a novel paradigm born from the limitations of…
CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions and Perception
arXiv:2608.08392v1 Announce Type: new Abstract: Large language models are increasingly deployed as autonomous agents that interact with the web through…
LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving
arXiv:2608.08382v1 Announce Type: new Abstract: As LLM inference shifts to multi-tenant GPU clusters, co-batching improves throughput but obscures…
Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning
arXiv:2608.08303v1 Announce Type: new Abstract: Agentic skills improve large language model (LLM) agents by encoding reusable procedures for complex…
Mitigating Over-Personalization in LLMs via Structured Memory
arXiv:2608.08300v1 Announce Type: new Abstract: Conversational assistants increasingly rely on persistent long-term memory to personalize responses across…
