arXiv:2609.00718v1 Announce Type: new Abstract: Many automobile and mobility companies deploy learned driving policies on embedded computers with limited…
Category: cs.AI updates on arXiv.org
ChatDev 2.0: A No-Code Multi-Agent Platform for Developing Everything
arXiv:2609.00714v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems (MAS) have shown strong potential for solving complex…
Triple-Bottom-Line Sustainability of Language Models for Edge AI: A Comparison Between SLMs and Quantized LLMs
arXiv:2609.00665v1 Announce Type: new Abstract: Edge-AI model selection is commonly driven by one isolated metric – accuracy, latency, memory, energy, or…
Value Over Language Model: Detecting Original Contribution in Writing
arXiv:2609.00700v1 Announce Type: new Abstract: LLMs have been rapidly adopted across writing tasks, prompting the development of tools for detecting…
DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation
arXiv:2609.00646v1 Announce Type: new Abstract: Commercial short-drama production follows a multi-stage chain: script, storyboard, keyframe imagery,…
Drift-Aware LLM Routing with Sparse Contexts and Shared Budgets
arXiv:2609.00662v1 Announce Type: new Abstract: A multi-model language service must route each request while preserving workload-level budgets for…
Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
arXiv:2609.00652v1 Announce Type: new Abstract: Language model agents increasingly propose actions, observe external feedback, and explain their own…
SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task
arXiv:2609.00654v1 Announce Type: new Abstract: We describe the SciTrue team’s participation in both subtasks of the NTCIR-19 SciClaimEval…
REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows
arXiv:2609.00643v1 Announce Type: new Abstract: Agent revisions expose a fundamental correctness–efficiency trade-off during concurrent execution.…
Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs
arXiv:2609.00575v1 Announce Type: new Abstract: Mixture-of-experts (MoE) architectures scale large language models efficiently, but they demand massive…
