arXiv:2608.17341v1 Announce Type: new Abstract: AI planning is concerned with finding a sequence of actions that achieves a specified goal. It relies on…
Category: AI
Task-Aware Harness Provisioning for LLM Agents in Mission-Critical Infrastructure Operations
arXiv:2608.17433v1 Announce Type: new Abstract: LLM agents have been widely adopted to operate mission-critical infrastructure (MCI). These agents…
LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents
arXiv:2608.17393v1 Announce Type: new Abstract: Reinforcement learning for coding agents increasingly relies on long-running agent harnesses to manage…
Depth Enables Local Entropy: Quadratic Depth Dependence in Deep Variation-Norm ReLU Regression
arXiv:2608.17434v1 Announce Type: new Abstract: We study Gaussian regression over the explicit vector-valued Parhi–Nowak deep-RBV^2 architecture with…
Amazon, which started off selling books, is destroying rare texts to train AI
Rare books are incredibly valuable for training LLMs, since these models have already trained on whatever’s available online.
Cognitive Graph Intelligence for Adaptive and Robust DDoS Attack Detection in Next Generation Networks
arXiv:2608.17352v1 Announce Type: new Abstract: Distributed Denial-of-Service (DDoS) attacks threaten network availability, requiring a cognitive…
Wuying-Browser-Agent: Real-World Centric Fundamental Long-Horizon Browser Agents
arXiv:2608.17319v1 Announce Type: new Abstract: Browser agents perform well on short, clean demonstrations, but real deployment is fundamentally…
TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration
arXiv:2608.17336v1 Announce Type: new Abstract: Long-context prefill in large language models (LLMs) incurs substantial computation and memory traffic…
LiveHouse-TS: An Open-world Living Benchmark for Time Series Foundation Models
arXiv:2608.17299v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently emerged as a highly promising paradigm for…
LLMs for Medical Consultation Are Evaluated Too Late: The Preformulation Gap
arXiv:2608.17330v1 Announce Type: new Abstract: Large language models for medical consultation are often evaluated after a clinical problem has already…
