arXiv:2609.11155v1 Announce Type: new Abstract: Multi-Agent Reinforcement Learning (MARL) has emerged as a pivotal paradigm for complex decision-making in…
Category: cs.AI updates on arXiv.org
Breaking Predictions Is Not Enough: Specified-Foil Counterfactuals for Temporal Graphs
arXiv:2609.11170v1 Announce Type: new Abstract: Temporal graph counterfactual explanations typically change past events to change or invalidate an…
The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems
arXiv:2609.11146v1 Announce Type: new Abstract: AI-generated text is flowing back into the training corpora of the next generation of models. Recursive…
Autonomous Chemical Mechanistic Discovery through Agentic Reasoning and Validation
arXiv:2609.11147v1 Announce Type: new Abstract: Unraveling reaction mechanisms is central to modern chemistry, yet automating these investigations remains…
Fork Where the Model Changes Its Mind: Belief-Shift Branching for Tree-Structured Reinforcement Learning
arXiv:2609.11061v1 Announce Type: new Abstract: Tree-structured rollouts give critic-free reinforcement learning with verifiable rewards (RLVR) step-level…
MOSAIC: Query-Aware Exploration Policy Adaptation for GraphRAG
arXiv:2609.11065v1 Announce Type: new Abstract: Graph Retrieval-Augmented Generation (GraphRAG) can connect evidence distributed across a corpus graph,…
KuaiRP Series Role-playing Models Technical Report
arXiv:2609.11127v1 Announce Type: new Abstract: This paper introduces the complete technical solution for the KuaiRP series of role-playing models. We aim…
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
arXiv:2609.11115v1 Announce Type: new Abstract: Benchmark researchers and developers of large language models (LLMs) and other AI systems need to find…
Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial Sentiment
arXiv:2609.11144v1 Announce Type: new Abstract: Financial NLP has a standard workflow: validate a sentiment tool against human labels, then trust it to…
Grounding Agent Memory: Environment-Probing Curation for Enterprise Agents
arXiv:2609.11060v1 Announce Type: new Abstract: Persistent memory is entering production-oriented agent platforms to help long-horizon agents accumulate…
