arXiv:2608.07107v2 Announce Type: replace Abstract: World models are increasingly used to support planning in agents by predicting how environment states…
Category: cs.AI updates on arXiv.org
Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability
arXiv:2608.11238v2 Announce Type: replace Abstract: Retrieval-augmented generation improves the factuality of large language models by grounding responses…
Fragility of Value under Imperfect Alignment
arXiv:2607.28881v4 Announce Type: replace Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that…
The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs
arXiv:2607.08734v2 Announce Type: replace Abstract: Post-Training Quantization has become widely used to compress large language models to make them…
Share the Judge, Learn the Deferral: Where Specialization Helps LLM Evaluation
arXiv:2607.27984v2 Announce Type: replace Abstract: Agentic systems generate outputs faster than human review. We contrast two LLM evaluator…
When Words Are Safe But Actions Kill: Probing Physical Jailbreak Beyond Textual Jailbreak in Hidden-State Risk Space
arXiv:2607.15218v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly serve as high-level planners for embodied agents, where…
WorldLines: Benchmarking and Modeling Long-Horizon Stateful Embodied Agents
arXiv:2606.18847v2 Announce Type: replace Abstract: To assist humans over extended periods in real homes, embodied agents must remember user routines,…
AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance
arXiv:2606.30949v2 Announce Type: replace Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world…
SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment
arXiv:2606.02530v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human values often degrades their general capabilities,…
Can LLMs Introspect? A Reality Check
arXiv:2605.26242v2 Announce Type: replace Abstract: Can large language models detect and report their own internal states? A number of recent studies have…
