arXiv:2609.13298v1 Announce Type: cross Abstract: Anomaly detection in structured images is challenging in small-data settings where deep learning…
Category: cs.AI updates on arXiv.org
LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents
arXiv:2609.13287v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) achieve high decoding efficiency through block-parallel,…
Adaptive Conformal Redistribution for Inter-class Transitional Uncertainty in Medical Image Classification
arXiv:2609.13303v1 Announce Type: cross Abstract: Medical image classification is frequently complicated by transitional categories whose feature…
IMM-based Multiple Object Tracking using a State Prediction Neural Network
arXiv:2609.13307v1 Announce Type: cross Abstract: Object tracking is essential for autonomous vehicles to avoid obstacles and plan routes. Radar maintains…
From Process Loss to Assembly Bonus: Human-Grounded Diagnosis of Multi-Agent LLM Collaboration
arXiv:2609.13261v1 Announce Type: cross Abstract: LLM agents are increasingly used for collaborative problem solving and human-group simulation. This…
BEACON: Behavior and Appearance Control for Subject-Specific Video Generation
arXiv:2609.13264v1 Announce Type: cross Abstract: Generating human-centric videos that preserve both visual identity and person-specific expressive…
Forward-Facing Near-Infrared Adds Little to Colour for Farm-Machinery Traversability: A Site-Disjoint Evaluation of Sensor-Dependent Spatial Leakage
arXiv:2609.13265v1 Announce Type: cross Abstract: Near-infrared (NIR) imaging does not consistently outperform standard color cameras for daytime…
Multimodal-Multiresolution Foundation Model for Lunar Remote Sensing
arXiv:2609.13283v1 Announce Type: cross Abstract: We present a multimodal foundation model for lunar remote sensing, pretrained from scratch on SomBench,…
Conflict-Predictive Variable Horizons in Multi-Drone Distributed Model Predictive Control
arXiv:2609.13270v1 Announce Type: cross Abstract: In distributed model predictive control for multi-drone collision avoidance, a fixed prediction horizon…
TryOnReward: Learning Foveated Consistency for Reinforcement Fine-Tuning of Virtual Try-On
arXiv:2609.13259v1 Announce Type: cross Abstract: Virtual Try-On (VTON) aims to dress a person with the reference garment, producing visually reasonable…
