arXiv:2608.22979v1 Announce Type: new Abstract: Deploying high-dimensional multimodal features in industrial recommender systems incurs substantial…
Category: cs.AI updates on arXiv.org
Toward Effective and Reliable LLM Agents via Dynamic Ontology
arXiv:2608.22974v1 Announce Type: new Abstract: Large language model (LLM) agents rely heavily on knowledge encoded in model parameters or presented as…
ParallelWorld: Test-Time Scaling for Embodied Reasoning
arXiv:2608.22971v1 Announce Type: new Abstract: Embodied Reasoning constitutes a fundamental capability of embodied intelligence, serving as the basis for…
CDEG: Learning Decision-Critical Evidence for Long-Horizon Diagnostic Agents
arXiv:2608.22899v1 Announce Type: new Abstract: Unlike static medical question answering, long-horizon diagnosis captures the sequential nature of…
Concepts for Securing Agentic AI Coding and the Terok Environment
arXiv:2608.22930v1 Announce Type: new Abstract: Agentic AI is a fascinating new tool for software development. It is a huge step forward compared to…
Proxy reliance in large language model decisions is uncalibrated to predictive evidence
arXiv:2608.22887v1 Announce Type: new Abstract: Large language models (LLMs) are entering decisions in triage and lending, where task-relevant inference…
Beyond Observed Auxiliary Relations: Environment-Conditioned Modeling for Multi-Behavior Recommendation
arXiv:2608.22920v1 Announce Type: new Abstract: Multi-behavior recommendation (MBR) leverages auxiliary behavioral signals, such as clicks and…
What Process Evaluation of Coding Agents Actually Measures: Action, Task, and Step Are Three Different Levels
arXiv:2608.22960v1 Announce Type: new Abstract: Coding agents are increasingly evaluated not only by whether they solve a task, but also by how they…
Beyond the Harness: End-to-End Optimization of Context Artifacts for Enterprise Text-to-SQL
arXiv:2608.22830v1 Announce Type: new Abstract: Deploying LLMs for enterprise Text-to-SQL is bottlenecked less by the model than by what context reaches…
Let the Bullets Fly: Multimodal Fake News Detection with Temporal-Aligned Generative Danmaku
arXiv:2608.22832v1 Announce Type: new Abstract: The social interactions among crowds via \textit{Danmaku} (a.k.a., bullet comments) on modern multimedia…
