arXiv:2608.17928v1 Announce Type: cross Abstract: In the Lifelong Multi-Agent Path Finding (L-MAPF) problem, agents must repeatedly move from one…
Category: cs.AI updates on arXiv.org
Efficient RLVR Scheduling via Graph-Structured Online Difficulty Estimation
arXiv:2608.17941v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large…
Analysis of Types of Inquiries in Student-AI Interaction: A case study of two CS2 tasks
arXiv:2608.17919v1 Announce Type: cross Abstract: Background and Context: Question and inquiry are integral parts of knowledge seeking and learning.…
AdaLens: Interactive Storyline for Monitoring and Steering Long-Running Agentic Data Analysis
arXiv:2608.17834v1 Announce Type: cross Abstract: Large language models are pushing data science toward increasingly autonomous and agentic workflows,…
Comparative Study of Out-of-the-Box Technology for Automatic Target Detection and Recognition
arXiv:2608.17917v1 Announce Type: cross Abstract: Automatic Target Detection and Recognition (ATD/R) is critical for military decision support and…
Encoded but Not Actionable: Auditing the Decode-Generate-Steer Gap in Frozen LLMs for Geometric Constraints
arXiv:2608.17843v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance on structured reasoning tasks, but…
The Model’s Tell: Measuring Context-Leakage Attack Signals with Behavior Gauges
arXiv:2608.17829v1 Announce Type: cross Abstract: LLMs increasingly rely on external contexts, such as pre-defined system prompts or retrieved documents,…
BEAR-Bench: A Bilingual Enterprise and Academic Reasoning Benchmark for Multimodal Models
arXiv:2608.17895v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) have made significant strides in visual comprehension,…
Interpretable Humans, Alien LLMs: Expert Analysis of Latent Structures in Assessment Responses
arXiv:2608.17810v1 Announce Type: cross Abstract: The evaluation of large language models (LLMs) relies heavily on human-designed assessments, implicitly…
Learnware for CSI Feedback: Scene-specific Small Models Can Do Big
arXiv:2608.17760v1 Announce Type: cross Abstract: Intelligent channel state information (CSI) feedback is essential for realizing the high capacity and…
