arXiv:2607.21617v2 Announce Type: replace Abstract: Vision Language Models (VLMs) are increasingly used in place of traditional OCR pipelines for document…
Tag: cs.AI updates on arXiv.org
Pailitao-MMSearch: Building Native E-Commerce Multimodal Search Foundation
arXiv:2607.17499v2 Announce Type: replace Abstract: The evolution of e-commerce has fundamentally transformed how users search for products, shifting from…
SeerGuard: A Safety Framework for Mobile GUI Agents via World Model Prediction
arXiv:2607.15550v3 Announce Type: replace Abstract: Mobile graphical user interface (GUI) agents have demonstrated remarkable capabilities in automating…
Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions
arXiv:2607.12406v2 Announce Type: replace Abstract: The capability of LLM agents to function as the “brain” of a system fundamentally expands the scope…
CHASE: Cache-Hole-Adapted Skip Exit for Looped State-Space Language Models
arXiv:2607.10110v2 Announce Type: replace Abstract: Recent work on looped language models suggests that many reasoning problems benefit from greater…
TopoBrick: Agentic Topology Sampling of Exogenous Variables for Zero-Shot Building IoT Forecasting
arXiv:2607.06349v2 Announce Type: replace Abstract: Building sensors are embedded in physical topology, spatial hierarchy, and operational context, yet…
EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent
arXiv:2606.17698v3 Announce Type: replace Abstract: As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper’s…
Heaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language Models
arXiv:2607.04562v3 Announce Type: replace Abstract: Large language models (LLMs) generate fluent outputs that can be wrong. Unlike humans, who often…
Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability?
arXiv:2606.24026v2 Announce Type: replace Abstract: Mechanistic interpretability has made substantial progress in automatically localizing circuits, but…
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation
arXiv:2605.29829v2 Announce Type: replace Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems…
