arXiv:2607.24780v2 Announce Type: replace Abstract: Fixed benchmarks are costly to renew and cannot adapt their questions to model-specific failures. We…
Category: cs.AI updates on arXiv.org
Adaptive Graph-of-Islands Evolution for Automatic Feature Engineering with LLMs
arXiv:2607.23286v2 Announce Type: replace Abstract: Automatic feature engineering (AutoFE) for tabular data requires discovering informative…
Reinforcement Learning for Heterogeneous Sensor Selection in Maritime Surveillance
arXiv:2607.22667v2 Announce Type: replace Abstract: This paper presents an information-gain-guided reinforcement-learning sensor-selection framework for…
PIE-APT: Abductive Planning over Temporal Dynamic Knowledge Graphs via Incremental Reasoning
arXiv:2607.27287v2 Announce Type: replace Abstract: Planning over Temporal Dynamic Knowledge Graphs (TDKGs) presents theoretical challenges in open-world…
CUSUM-Shaped Inference-Time Monitoring and Targeted Re-Decoding for Quantized Small Language Model Reasoning
arXiv:2607.20129v2 Announce Type: replace Abstract: Quantized small reasoning models can enter repetitive or otherwise unproductive trajectories, yet…
Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models
arXiv:2607.21617v2 Announce Type: replace Abstract: Vision Language Models (VLMs) are increasingly used in place of traditional OCR pipelines for document…
Pailitao-MMSearch: Building Native E-Commerce Multimodal Search Foundation
arXiv:2607.17499v2 Announce Type: replace Abstract: The evolution of e-commerce has fundamentally transformed how users search for products, shifting from…
SeerGuard: A Safety Framework for Mobile GUI Agents via World Model Prediction
arXiv:2607.15550v3 Announce Type: replace Abstract: Mobile graphical user interface (GUI) agents have demonstrated remarkable capabilities in automating…
Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions
arXiv:2607.12406v2 Announce Type: replace Abstract: The capability of LLM agents to function as the “brain” of a system fundamentally expands the scope…
CHASE: Cache-Hole-Adapted Skip Exit for Looped State-Space Language Models
arXiv:2607.10110v2 Announce Type: replace Abstract: Recent work on looped language models suggests that many reasoning problems benefit from greater…
