arXiv:2512.15662v4 Announce Type: replace Abstract: Human beings solve complex problems through critical thinking, where reasoning and evaluation are…
Category: cs.AI updates on arXiv.org
Achieving Olympiad-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning
arXiv:2512.10534v4 Announce Type: replace Abstract: Large language model (LLM) agents exhibit strong mathematical problem-solving abilities and can even…
What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?
arXiv:2512.24497v4 Announce Type: replace Abstract: A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks…
Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM Agents
arXiv:2601.07468v2 Announce Type: replace Abstract: Memory enables Large Language Model (LLM) agents to perceive, store, and use information from past…
Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework
arXiv:2609.02861v1 Announce Type: cross Abstract: Autonomous robots powered by deep learning face a fundamental auditability challenge: when incidents…
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning
arXiv:2505.15276v2 Announce Type: replace Abstract: Large reasoning models (LRMs) have achieved remarkable success on complex tasks, yet their tendency to…
Modeling and Optimizing User Preferences in AI Copilots: A Comprehensive Survey and Taxonomy
arXiv:2505.21907v3 Announce Type: replace Abstract: AI copilots represent a new generation of AI-powered systems designed to assist users, particularly…
Post-Training Language Models for Gold-Medal Performance in Coding Competitions
arXiv:2609.02849v1 Announce Type: cross Abstract: Competitive programming has become a key test of large language model reasoning, with international…
AI Mathematician: Towards Fully Automated Frontier Mathematical Research
arXiv:2505.22451v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have made significant progress in mathematical capabilities in recent…
From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution
arXiv:2609.02771v1 Announce Type: cross Abstract: Training data attribution (TDA) aims to identify training examples that shape model behavior, but its…
