arXiv:2608.17605v2 Announce Type: replace-cross Abstract: Conversational AI is moving beyond isolated text prompts toward sustained, multimodal…
Category: cs.AI updates on arXiv.org
Self-Explanation Tutor for Active Study of CS1 Worked Examples
arXiv:2608.25180v2 Announce Type: replace-cross Abstract: Worked examples are an important part of introductory programming, but reading their expert…
Modeling Human Behavior with Type Vectors Using AI
arXiv:2608.18265v4 Announce Type: replace-cross Abstract: We introduce a general, easy-to-implement AI-based modeling technique for analyzing human…
Sixteen models, fewer than two voices: measuring ensemble dispersion where no answer is uniquely correct
arXiv:2608.00285v2 Announce Type: replace-cross Abstract: Sixteen language models drawn from ten families produced, on average, the semantic diversity of…
Prompt-Driven Exploration: Language as an Exploration Space for VLA Reinforcement Learning
arXiv:2607.08837v4 Announce Type: replace-cross Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the…
Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification
arXiv:2607.13571v2 Announce Type: replace-cross Abstract: Sound event detection relies on frame-level strong labels whose annotation is expensive. Active…
Teaching LLMs to Self-Evolve: Cultivating Core Meta-Skills with Reinforcement Learning
arXiv:2607.21971v2 Announce Type: replace-cross Abstract: Test-time scaling through iterative self-evolution with environment feedback, as demonstrated by…
Self-Reference in Large Language Models: The Introspection Threshold for Recursive Self-Improvement
arXiv:2607.04277v2 Announce Type: replace-cross Abstract: The pursuit of self-evolving AI raises a critical question: when is autonomous self-improvement…
Git-Assistant: Planning-Based Support for Updating Git Repositories
arXiv:2607.09224v3 Announce Type: replace-cross Abstract: Version control systems are essential for collaborative software development, yet tools like git…
Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications
arXiv:2607.00442v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for quadruped locomotion commonly depends on fixed, hand-crafted,…
