arXiv:2609.00662v1 Announce Type: new Abstract: A multi-model language service must route each request while preserving workload-level budgets for…
Tag: cs.AI updates on arXiv.org
Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
arXiv:2609.00652v1 Announce Type: new Abstract: Language model agents increasingly propose actions, observe external feedback, and explain their own…
SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task
arXiv:2609.00654v1 Announce Type: new Abstract: We describe the SciTrue team’s participation in both subtasks of the NTCIR-19 SciClaimEval…
REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows
arXiv:2609.00643v1 Announce Type: new Abstract: Agent revisions expose a fundamental correctness–efficiency trade-off during concurrent execution.…
Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs
arXiv:2609.00575v1 Announce Type: new Abstract: Mixture-of-experts (MoE) architectures scale large language models efficiently, but they demand massive…
Consistency Without Alignment: Item-Sensitive Language Models Indistinguishable From Random
arXiv:2609.00576v1 Announce Type: new Abstract: Item-sensitivity, defined as whether a model’s choice depends on the specific input rather than on its own…
Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs
arXiv:2609.00621v1 Announce Type: new Abstract: Prompt optimization can improve multi-agent LLM systems, but the prompts being optimized often serve two…
Socrates went Nuclear: Comparing Interaction Strategies for AI systems in a Learning Context using Brain Sensing
arXiv:2609.00584v1 Announce Type: new Abstract: Does unrestricted AI access bypass the cognitive effort required for learning, or does it streamline…
Same Request, Different Boundary: Evaluating Cybersecurity Assistance across Conversational Contexts
arXiv:2609.00578v1 Announce Type: new Abstract: Large Language Models (LLMs) can solve complex problems, but their misuse in high-risk domains can lead to…
ISO-RAG: Isoperimetric Noise Control for Retrieval-Augmented Generation
arXiv:2609.00513v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) mitigates large language models (LLMs) hallucinations, yet…
