arXiv:2608.24063v1 Announce Type: cross Abstract: While Vision Large Language Models (VLLMs) have achieved remarkable success in multimodal reasoning,…
Category: cs.AI updates on arXiv.org
ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning
arXiv:2608.24033v1 Announce Type: cross Abstract: Time series classification underpins applications in healthcare, sensing, and industrial monitoring.…
Don’t Just Listen, Try Planning: Graph-based Retrieval-Generation Agent for Long-form Audio Meeting Understanding
arXiv:2608.24048v1 Announce Type: cross Abstract: While long-form audio meeting understanding (LAMU) is garnering growing attention, task-specific…
WebMCP-Phalanx: Enforcing and Characterizing Trust Boundaries for Browser-Integrated LLM Agents
arXiv:2608.24017v1 Announce Type: cross Abstract: The emerging W3C WebMCP proposal enables LLM agents to invoke tools exposed by web pages. In multi-party…
IterCAD: Iterative Program Repair for CAD Code Generation from Orthographic Views
arXiv:2608.24020v1 Announce Type: cross Abstract: Generating executable parametric CAD code from dimension-annotated orthographic drawings is a…
What Guides the Agent? Adjudicating Unauthorized Behavior via Localizing Behavior-Guiding Instructions
arXiv:2608.24022v1 Announce Type: cross Abstract: LLM agents integrated with external resources gain complex task capabilities, yet the unified…
Hybrid Semantic Tool Discovery for Enterprise MCP Gateway: Architecture and Implementation
arXiv:2608.23992v1 Announce Type: cross Abstract: Large language model (LLM) agents invoke external tools to retrieve and reason over information beyond…
SAGE: From Direct Answering to Evidence-Grounded Inference for Chinese Ancient Document Understanding
arXiv:2608.24011v1 Announce Type: cross Abstract: Chinese ancient document understanding demands complex visual, linguistic, and historical reasoning.…
The Empire, Long Divided, Must Unite: Architectural Convergence in Three LLM Agent Harnesses
arXiv:2608.23953v1 Announce Type: cross Abstract: An agent harness is what turns a language model into an autonomous agent: the surrounding code that…
Evaluating Language Models on Cross-Language Code Functional Equivalence
arXiv:2608.23961v1 Announce Type: cross Abstract: Background: Large Language Models (LLMs) have demonstrated strong performance across a variety of…
