arXiv:2609.08188v1 Announce Type: new Abstract: Vision-language models (VLMs) augmented with retrieval-augmented generation (RAG) benefit from access to…
Tag: cs.AI updates on arXiv.org
Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collapse in Large Reasoning Models
arXiv:2609.08186v1 Announce Type: new Abstract: The emergence of Chain-of-Thought (CoT) has established a robust foundation for Large Reasoning Models…
Less Is Personal: Learning Minimal Sufficient User Profiles for Personalized Language Models
arXiv:2609.08180v1 Announce Type: new Abstract: Retrieval-augmented personalization enables large language models to produce more accurate and…
OntologyBench: Can Dense Retrieval Satisfy Structured Biomedical Constraints?
arXiv:2609.08174v1 Announce Type: new Abstract: We introduce OntologyBench, a tiered biomedical retrieval benchmark comprising 471,854 training and…
WorldAgen: Unified State-Action Prediction with Test-Time World Model Training
arXiv:2609.08162v1 Announce Type: new Abstract: How can vision-language-action (VLA) models adapt to new environments where world dynamics shift? While…
SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
arXiv:2609.08126v1 Announce Type: new Abstract: We study scheming in LLM agents, in which agents covertly pursue misaligned goals. Our focus is to…
Router Prior Bias: Preserving Base Routing Structure in MoE Post-Training
arXiv:2609.08115v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) pretraining relies on an auxiliary load-balancing loss (LBL) to drive per-expert…
SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents
arXiv:2609.08149v1 Announce Type: new Abstract: SWE-Bench Pro has emerged as a standard benchmark for evaluating software engineering agents on…
Key Path Identification for Resolving Knowledge Conflicts via SAE-based Steering
arXiv:2609.08173v1 Announce Type: new Abstract: Sparse autoencoder (SAE)-based steering has been widely used to address knowledge conflicts by guiding…
RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode Recognition in Older Adults and Clinical Cohorts
arXiv:2609.08090v2 Announce Type: new Abstract: Assistive devices for people with mobility impairments, such as powered exoskeletons, rely on accurate…
