arXiv:2608.15018v2 Announce Type: replace Abstract: Deploying large language models (LLMs) for inference on edge devices is challenging due to severe…
Category: cs.AI updates on arXiv.org
Reconstruction: A Blind Benchmark for Recovering Research Ideas from Pre-Publication Bibliographies
arXiv:2608.16645v2 Announce Type: replace Abstract: Can a language model recover the true research idea of a published paper when given only that paper’s…
GRIP: Grounded Reasoning via Information-Restricted Premises
arXiv:2608.16776v2 Announce Type: replace Abstract: High-capacity encoders in retrieval-augmented generation (RAG) can let the query dominate the latent…
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks
arXiv:2608.03502v2 Announce Type: replace Abstract: Large Language Models (LLMs) have recently shown strong capabilities in reasoning, planning, and…
Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations
arXiv:2607.20379v2 Announce Type: replace Abstract: Natural-language autoencoders score explanations of hidden activations by reconstruction. An…
Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting
arXiv:2607.26643v2 Announce Type: replace Abstract: Enabling large language model (LLM) agents to accumulate and reuse experience from past interactions…
BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding
arXiv:2608.04156v2 Announce Type: replace Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it…
Fragility of Value under Imperfect Alignment
arXiv:2607.28881v3 Announce Type: replace Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that…
ContextSniper: AntTrail’s Token-Efficient Code Memory for Repository-Level Program Repair
arXiv:2607.01916v5 Announce Type: replace Abstract: Large language model agents can repair real repository issues, but they often spend large context…
ReasFlow: Assisting Reasoning-Centric Scientific Discovery in Applied Mathematics via a Knowledge-Based Multi-Agent System
arXiv:2607.14178v3 Announce Type: replace Abstract: Recent advances in Large Language Models have fueled autonomous AI agents capable of tackling complex…
