arXiv:2609.17479v1 Announce Type: cross Abstract: Despite the rapid uptake of black-box object detectors in marine mammal research and monitoring,…
Category: cs.AI updates on arXiv.org
Evaluating Verified Autonomy in Quantum Engineering
arXiv:2609.17439v1 Announce Type: cross Abstract: Reliable quantum engineering is essential for turning quantum phenomena into practical technologies. As…
CareMirror: Bringing Caregiver Wellbeing into the Dementia Care Ecosystem
arXiv:2609.17434v1 Announce Type: cross Abstract: Family caregivers of people living with dementia shoulder emotional and practical responsibilities, yet…
Coupled Calibration and Learning: Mitigating Teacher Bias in LLM Distillation without Target-Domain Reward Feedback
arXiv:2609.17474v1 Announce Type: cross Abstract: Large language model (LLM) distillation aims to transfer the capabilities of a powerful teacher to a…
Where Should a Document Live: Context, Representations, or Parameters?
arXiv:2609.17346v1 Announce Type: cross Abstract: To answer questions outside of their pre-training data, large language models (LLMs) need access to new…
Learning-Guided Planning in Large Dynamic Action Spaces: Budgeted Tree Search for One-to-Many Mobile Charging
arXiv:2609.17429v1 Announce Type: cross Abstract: Many learned sequential decision systems map the current state directly to an action. That shortcut…
Tracking the Unseen: An Occlusion-Robust Framework for Target Tracking Under Full and Long-Term Occlusion
arXiv:2609.17427v1 Announce Type: cross Abstract: Real-time multi-object tracking systems remain highly vulnerable to full and long-term occlusion, where…
CTAN: Cycle-Temporal Attention Network for Embodied Audio-Visual Navigation
arXiv:2609.17420v1 Announce Type: cross Abstract: Audio-visual embodied navigation equips robots with the capability to infer the locations of sound…
Coding Agents Have Converged: Why the SWE-bench Leaderboard Can No Longer Order Its Top Entries, and What to Measure Instead
arXiv:2609.17394v1 Announce Type: cross Abstract: Small differences on coding-agent leaderboards are often read as an ordering of systems. We audit…
Easy to Catch a Liar, Hard to Clear an Honest One: Language Models Diagnosing a Corrupted Reward Channel from a Verified Record
arXiv:2609.17226v1 Announce Type: cross Abstract: An agent that learns from rewards has to trust whatever reports those rewards. When the reports suddenly…
