arXiv:2608.08244v2 Announce Type: replace Abstract: General-purpose wearable foundation models are pretrained on broad sensor streams and populations, but…
Tag: cs.AI updates on arXiv.org
SKILL.state: Scalable Long-Horizon Agent Skills
arXiv:2608.26263v3 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly act as autonomous agents executing complex, long-running…
Nova: An End-to-End MLIR Compiler for Deep Learning
arXiv:2608.00029v3 Announce Type: replace Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level…
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
arXiv:2608.08389v2 Announce Type: replace Abstract: Long-horizon research agents solve open-ended tasks through iterative retrieval, aggregation, and…
Aletheia: An Offline-First Clinical Decision Support System for Differential Diagnosis in Low-Resource Healthcare Settings
arXiv:2607.24814v2 Announce Type: replace Abstract: Access to specialist clinical expertise remains severely limited across sub-Saharan Africa, where…
LivingArena: Do LLMs Know What Other LLMs Don’t? Peer-Probing as Scalable Evaluation
arXiv:2607.24780v2 Announce Type: replace Abstract: Fixed benchmarks are costly to renew and cannot adapt their questions to model-specific failures. We…
Adaptive Graph-of-Islands Evolution for Automatic Feature Engineering with LLMs
arXiv:2607.23286v2 Announce Type: replace Abstract: Automatic feature engineering (AutoFE) for tabular data requires discovering informative…
Reinforcement Learning for Heterogeneous Sensor Selection in Maritime Surveillance
arXiv:2607.22667v2 Announce Type: replace Abstract: This paper presents an information-gain-guided reinforcement-learning sensor-selection framework for…
PIE-APT: Abductive Planning over Temporal Dynamic Knowledge Graphs via Incremental Reasoning
arXiv:2607.27287v2 Announce Type: replace Abstract: Planning over Temporal Dynamic Knowledge Graphs (TDKGs) presents theoretical challenges in open-world…
CUSUM-Shaped Inference-Time Monitoring and Targeted Re-Decoding for Quantized Small Language Model Reasoning
arXiv:2607.20129v2 Announce Type: replace Abstract: Quantized small reasoning models can enter repetitive or otherwise unproductive trajectories, yet…
