arXiv:2608.17393v1 Announce Type: new Abstract: Reinforcement learning for coding agents increasingly relies on long-running agent harnesses to manage…
Category: cs.AI updates on arXiv.org
Depth Enables Local Entropy: Quadratic Depth Dependence in Deep Variation-Norm ReLU Regression
arXiv:2608.17434v1 Announce Type: new Abstract: We study Gaussian regression over the explicit vector-valued Parhi–Nowak deep-RBV^2 architecture with…
Cognitive Graph Intelligence for Adaptive and Robust DDoS Attack Detection in Next Generation Networks
arXiv:2608.17352v1 Announce Type: new Abstract: Distributed Denial-of-Service (DDoS) attacks threaten network availability, requiring a cognitive…
Wuying-Browser-Agent: Real-World Centric Fundamental Long-Horizon Browser Agents
arXiv:2608.17319v1 Announce Type: new Abstract: Browser agents perform well on short, clean demonstrations, but real deployment is fundamentally…
TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration
arXiv:2608.17336v1 Announce Type: new Abstract: Long-context prefill in large language models (LLMs) incurs substantial computation and memory traffic…
LiveHouse-TS: An Open-world Living Benchmark for Time Series Foundation Models
arXiv:2608.17299v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently emerged as a highly promising paradigm for…
LLMs for Medical Consultation Are Evaluated Too Late: The Preformulation Gap
arXiv:2608.17330v1 Announce Type: new Abstract: Large language models for medical consultation are often evaluated after a clinical problem has already…
SignalReasoner: Assessing the Upper Bound of 3B Models for Signal Mathematical Reasoning
arXiv:2608.17301v1 Announce Type: new Abstract: Post-training with supervised chain-of-thought fine-tuning and reinforcement learning from verifiable…
ASI-Bench: At the Dawn of Artificial Superintelligence
arXiv:2608.17271v1 Announce Type: new Abstract: Artificial superintelligence (ASI) requires AI to move beyond mastering existing knowledge toward…
PlanPO: Group Planning-Aware Policy Optimization for Multi-Turn Agentic LLMs
arXiv:2608.17289v1 Announce Type: new Abstract: Group-relative policy optimization has emerged as a key paradigm for training agentic large language…
