arXiv:2406.09770v2 Announce Type: replace-cross Abstract: Solving multi-objective optimization problems for large deep neural networks is a challenging…
Category: cs.AI updates on arXiv.org
Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Caption Retrieval
arXiv:2608.18521v2 Announce Type: replace Abstract: Dense-caption retrieval has recently been improved by introducing segmentation, edge maps,…
Teacher-free Latent Self-distillation and Class-separable Representations for Lightweight IoT Attack Detection
arXiv:2403.15509v3 Announce Type: replace-cross Abstract: Knowledge distillation (KD) has been widely used to improve lightweight AI models by…
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
arXiv:2608.07267v2 Announce Type: replace Abstract: Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models…
LongRCA Bench: Diagnosing Responsible Roles and Root Causes in Long-Horizon Agent Failures
arXiv:2608.15242v2 Announce Type: replace Abstract: When a long-horizon agent execution fails, outcome-level evaluation reveals the unsuccessful result…
Graph Surgery and the Do-Operator: A Precise Correspondence for Acyclic Structural Causal Models
arXiv:2608.17634v2 Announce Type: replace Abstract: The $\operatorname{do}$-operator is described graphically by deleting arrows into its targets and…
Andy: A Mathematical Agent for Rigorous Proof and Autonomous Research
arXiv:2608.15052v2 Announce Type: replace Abstract: Andy is an autonomous mathematical research agent that turns a mathematical problem into a traceable…
GENCO – A Unified Neural Solver Embedded in a Development Framework for Steady-State Grid Analysis
arXiv:2608.09921v2 Announce Type: replace Abstract: Foundation models are transforming business workflows and boosting productivity, yet they remain…
Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Personalzied Financial Agents
arXiv:2608.06108v2 Announce Type: replace Abstract: Investment competence is inherently personalized: the same market evidence can justify different…
CADRE: Stable, Parameter Efficient Adaptation of Medical Vision Language Models with Bounded Forgetting and Prior Drift
arXiv:2606.23487v2 Announce Type: replace Abstract: Medical vision-language models (VLMs) such as BiomedCLIP generalize broadly, but adapting them to a…
