arXiv:2609.00935v1 Announce Type: cross Abstract: Deep Research agents tackle knowledge-intensive tasks through multi-round retrieval and…
Category: cs.AI updates on arXiv.org
Benchmarking Vision-Language Models for Automated Pathology Diagnosis and Report Generation
arXiv:2609.00866v1 Announce Type: cross Abstract: The rapid advancement of vision-language models (VLMs) has accelerated progress in computational…
Beyond the Image Plane: World-Grounded Queries for Multi-Object Tracking
arXiv:2609.00924v1 Announce Type: cross Abstract: Monocular videos record 3D scenes as sequences of 2D image-plane projections, obscuring depth and…
Context-Grounding Gains Are Mediated by Pre-existing Machinery: Auditing GRPO, SFT, and DPO
arXiv:2609.00925v1 Announce Type: cross Abstract: Language models can ignore prompt evidence when it conflicts with memorized knowledge. Post-training can…
ADGNet: Asymmetric Dual-text Guided Network for Infrared Small Target Detection
arXiv:2609.00853v1 Announce Type: cross Abstract: InfRared Small Target Detection (IRSTD) is a challenging task. Relying solely on pixel-level…
Replacing Training with Memory: Listwise Selection for Text-to-SQL
arXiv:2609.00834v1 Announce Type: cross Abstract: Modern Text-to-SQL systems often follow generate-execute-select pipelines, generating multiple candidate…
Probabilistic Model Checking of Autoregressive Neural Sequence Models
arXiv:2609.00838v1 Announce Type: cross Abstract: Test-set accuracy is silent on two issues that matter when deploying autoregressive neural sequence…
Does Fault Localization Beat a Fresh Attempt? A Placebo-Controlled Study of Test-Guided Code Repair
arXiv:2609.00854v1 Announce Type: cross Abstract: Fault localization can focus a code model’s repair on the statements a failing test implicates, but a…
A Checklist to assess the energy and carbon impacts of ML/AI applications in Earth System Modeling
arXiv:2609.00847v1 Announce Type: cross Abstract: As machine learning and artificial intelligence find their way into nearly every aspect of climate,…
Visual Attention Faithfulness in Vision-Language Models is Heterogeneous
arXiv:2609.00830v1 Announce Type: cross Abstract: Whether attention weights faithfully reflect model reasoning has been actively debated in NLP, yet this…
