arXiv:2606.08093v2 Announce Type: replace Abstract: Pathology is the cornerstone of modern medicine, where accurate decision-making relies heavily on…
Category: cs.AI updates on arXiv.org
Thinking Before Retrieving: Robust Zero-Shot Composed Image Retrieval via Strategic Planning and Self-Criticism
arXiv:2606.31222v2 Announce Type: replace Abstract: Composed image retrieval requires identifying a target image from a gallery by integrating a reference…
ChatPlanner: A Large Language Model Framework for Personalized Public Transit Routing
arXiv:2606.15315v2 Announce Type: replace Abstract: Personalized public transit routing in public transit systems remains challenging due to the…
LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models
arXiv:2605.09948v2 Announce Type: replace Abstract: Current Vision-Language-Action (VLA) models typically treat the deepest representation of a…
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
arXiv:2604.08525v3 Announce Type: replace Abstract: Large language models (LLMs) are trained to align with user preferences through methods like…
ScreenSearch: Uncertainty-Aware OS Exploration
arXiv:2605.16024v2 Announce Type: replace Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to…
Pander Score: A Continuous Measure of Sycophancy as Epistemic Deference
arXiv:2606.07897v2 Announce Type: replace Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user.…
Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs
arXiv:2604.12616v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) expand the attack surface of safety-aligned systems by coupling visual…
Planning under Distribution Shifts with Causal POMDPs
arXiv:2602.23545v3 Announce Type: replace Abstract: In the real world, planning is often challenged by distribution shifts. As such, a model of the…
Does Unification Come at a Cost? Uni-SafeBench: A Safety Benchmark for Unified Multimodal Large Models
arXiv:2604.00547v2 Announce Type: replace Abstract: Unified Multimodal Large Models (UMLMs) integrate understanding and generation capabilities within a…
