arXiv:2608.10526v1 Announce Type: cross Abstract: Motivated by decentralized applications, we study cooperative multi-agent bandits in continuous…
Category: cs.AI updates on arXiv.org
Rethinking Text-Based Image Retrieval in Specific Domain
arXiv:2608.10524v1 Announce Type: cross Abstract: Driven by the rapid advancement of vision-language representation learning, Text-based Image Retrieval…
Unlocking the Power of Medical Tabular Data via Semantic-Aware Multimodal Pre-training
arXiv:2608.10522v1 Announce Type: cross Abstract: While vision-language models dominate medical representation learning, unstructured text lacks the…
Critic-Free Pretraining for Efficient Online Reinforcement Learning Fine-Tuning
arXiv:2608.10473v1 Announce Type: cross Abstract: Offline-to-online (O2O) reinforcement learning aims to leverage policies pretrained on static datasets…
SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning
arXiv:2608.10513v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) remain vulnerable to jailbreak attacks that exploit visual inputs…
Lost in Reconstruction: Aligning Action Representations with Language in Vision-Language-Action Models
arXiv:2608.10484v1 Announce Type: cross Abstract: Action verbs describe not only the physical outcomes of actions, but also how those actions are…
Exploration-Driven Personalized Federated Reinforcement Learning via Intrinsic Motivation
arXiv:2608.10499v1 Announce Type: cross Abstract: Personalized Federated Reinforcement Learning (PFRL) takes a decentralized approach to storing and…
Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging
arXiv:2608.10447v1 Announce Type: cross Abstract: Large language model-based recommender systems are increasingly adopting slow-thinking models that…
From Reasoning Depth to Reasoning Breadth: Evaluating Multi-Point Associative Reasoning in Large Language Models
arXiv:2608.10444v1 Announce Type: cross Abstract: Large language models (LLMs) have made substantial progress on reasoning tasks that require increasingly…
FUSE: Frame-Unified Stress Estimation from Facial Video
arXiv:2608.10442v1 Announce Type: cross Abstract: Automatic stress detection from facial video offers a practical path to non-intrusive affect monitoring,…
