arXiv:2511.08873v3 Announce Type: replace Abstract: Large language models (LLMs) are shifting from answer providers to intelligent tutors in educational…
Category: cs.AI updates on arXiv.org
ReflCtrl: Controlling LLM Reflection Efficiently via Representation Engineering
arXiv:2512.13979v2 Announce Type: replace Abstract: Large reasoning models achieve strong performance on diverse tasks by producing extended chains of…
Comparing Explanations is Not Enough, Explain the Change: New Standards are Needed to Explain Behavioral Shifts in Large Language Models
arXiv:2602.02304v3 Announce Type: replace Abstract: Large-scale foundation models exhibit behavioral shifts when subjected to interventions such as…
Illuminating the Three Dogmas of Reinforcement Learning under Evolutionary Light
arXiv:2507.11482v5 Announce Type: replace Abstract: Artificial learning systems are graduating from passive learners to increasingly autonomous agents,…
Efficient LLM Collaboration via Planning
arXiv:2506.11578v5 Announce Type: replace Abstract: Recently, large language models (LLMs) have demonstrated strong performance, ranging from simple to…
LEMMA-RCA: A Large Multi-modal Multi-domain Dataset for Root Cause Analysis
arXiv:2406.05375v4 Announce Type: replace Abstract: Root cause analysis (RCA) is crucial for enhancing the reliability and performance of complex systems.…
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
arXiv:2312.17535v2 Announce Type: replace Abstract: In the past two years, the outstanding performance of ChatGPT in multilingual and multitasking has led…
Adaptive GR(1) Specification Repair for Liveness-Preserving Shielding in Reinforcement Learning
arXiv:2511.02605v3 Announce Type: replace Abstract: Shielding is widely used to enforce safety in reinforcement learning (RL), ensuring that an agent’s…
Reading Is Not Using: Retrieval, Judgment, and the Design of AI Financial Research Workflows
arXiv:2608.24842v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as AI analysts to process financial disclosures…
Topology-Guided Modular Actor-Critic Learning for Continuous Systems under Temporal Objectives
arXiv:2304.10041v2 Announce Type: replace Abstract: This work investigates formal policy synthesis for continuous-state stochastic dynamic systems subject…
