arXiv:2609.02272v1 Announce Type: cross Abstract: Faithfully translating research papers into repository-level implementations remains challenging because…
Tag: cs.AI updates on arXiv.org
Auditory Illusion Benchmark for Large Audio Language Models
arXiv:2609.02277v1 Announce Type: cross Abstract: Perceptual illusions have long served as crucial probes into human cognition, revealing biases and…
Retrosynthesis of Synthetic Media for Explainable AI Provenance Forensics
arXiv:2609.02268v1 Announce Type: cross Abstract: With the rapid proliferation of generative models on Machine Learning as a Service (MLaaS) platforms,…
Do Large Language Models Capture the Diversity in their Training Data?
arXiv:2609.02275v1 Announce Type: cross Abstract: Large language models are trained to model conditional distributions over text, yet it remains…
InfraPatch: Cross-Task Targeted Grayscale Patch Attacks on Infrared-Adapted Vision-Language Models
arXiv:2609.02233v1 Announce Type: cross Abstract: Infrared vision-language models (IR-VLMs) have emerged as a promising paradigm for multimodal perception…
SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework
arXiv:2609.02203v1 Announce Type: cross Abstract: Time series representation learning (TSRL) has attracted growing research interests in recent years. Two…
Signal or Noise? Auditing Rotation-Induced Saliency Drift in Medical and Aerial Imaging
arXiv:2609.02224v1 Announce Type: cross Abstract: Post-hoc saliency maps such as Grad-CAM are increasingly used to audit why a deployed vision model made…
DiffuSearch: How Hybrid Trajectory Planning Benefits from Aligned Objectives in Diffusion and Action Space
arXiv:2609.02252v1 Announce Type: cross Abstract: In trajectory planning for autonomous driving, hybrid planning architectures are often realized as a…
SAUF-Net: Structure–Appearance Representation Learning with Uncertainty Feedback for Semi-Supervised Medical Image Segmentation
arXiv:2609.02247v1 Announce Type: cross Abstract: Semi-supervised learning has shown great potential for reducing annotation costs in medical image…
OmegaUse-SOP: SOP Engineering for Professional Computer Use from Human Demonstrations
arXiv:2609.02149v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly evolving from conversational assistants into agents…
