arXiv:2609.00700v1 Announce Type: new Abstract: LLMs have been rapidly adopted across writing tasks, prompting the development of tools for detecting…
Tag: AI
DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation
arXiv:2609.00646v1 Announce Type: new Abstract: Commercial short-drama production follows a multi-stage chain: script, storyboard, keyframe imagery,…
Drift-Aware LLM Routing with Sparse Contexts and Shared Budgets
arXiv:2609.00662v1 Announce Type: new Abstract: A multi-model language service must route each request while preserving workload-level budgets for…
Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
arXiv:2609.00652v1 Announce Type: new Abstract: Language model agents increasingly propose actions, observe external feedback, and explain their own…
SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task
arXiv:2609.00654v1 Announce Type: new Abstract: We describe the SciTrue team’s participation in both subtasks of the NTCIR-19 SciClaimEval…
REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows
arXiv:2609.00643v1 Announce Type: new Abstract: Agent revisions expose a fundamental correctness–efficiency trade-off during concurrent execution.…
Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs
arXiv:2609.00575v1 Announce Type: new Abstract: Mixture-of-experts (MoE) architectures scale large language models efficiently, but they demand massive…
Consistency Without Alignment: Item-Sensitive Language Models Indistinguishable From Random
arXiv:2609.00576v1 Announce Type: new Abstract: Item-sensitivity, defined as whether a model’s choice depends on the specific input rather than on its own…
Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs
arXiv:2609.00621v1 Announce Type: new Abstract: Prompt optimization can improve multi-agent LLM systems, but the prompts being optimized often serve two…
Socrates went Nuclear: Comparing Interaction Strategies for AI systems in a Learning Context using Brain Sensing
arXiv:2609.00584v1 Announce Type: new Abstract: Does unrestricted AI access bypass the cognitive effort required for learning, or does it streamline…
