arXiv:2608.24368v2 Announce Type: replace Abstract: Reliable multi-turn tool use requires an agent to preserve an evolving task state and ensure that each…
Category: cs.AI updates on arXiv.org
DataKernelBench: Can LLMs Optimize Database Queries on GPUs?
arXiv:2608.25061v2 Announce Type: replace-cross Abstract: GPUs increasingly accelerate database systems, but query-specific peak performance still often…
Unsupervised Post-Training of Foundation Models: A Survey
arXiv:2608.24982v2 Announce Type: replace-cross Abstract: Foundation-model post-training usually relies on human labels, preference data, stronger…
MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification
arXiv:2608.13463v2 Announce Type: replace-cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often…
X$^2$Localizer: Cross-grained Alignment for Progressive Cross-view Video Geo-localization
arXiv:2608.16658v2 Announce Type: replace-cross Abstract: Cross-view Video Geo-localization (CVG) aims to localize ground-view videos by retrieving their…
Pre-training Visual Dexterity in Simulation
arXiv:2608.15917v2 Announce Type: replace-cross Abstract: Large-scale pre-training has made robot policy fine-tuning increasingly data-efficient, but this…
Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules
arXiv:2608.22642v2 Announce Type: replace-cross Abstract: Despite recent advances in molecular foundation models, several limitations remain, such as…
Complexity Induction: Compositional Generalization via Structured Training Distortion
arXiv:2608.21464v2 Announce Type: replace-cross Abstract: We demonstrate that structured distortion of training data – which we term complexity induction…
A 12-CNOT Double Qubit Excitation Gate
arXiv:2608.11733v3 Announce Type: replace-cross Abstract: Effective implementation of high-level quantum gates is essential for practical quantum…
REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation
arXiv:2608.11698v3 Announce Type: replace-cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level…
