arXiv:2609.05539v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have recently made strong progress in vision-language…
Tag: AI
Subject-Relative Micro-Motion and Sleep Dynamics for Near-Infrared Video Sleep Staging
arXiv:2609.05550v1 Announce Type: cross Abstract: Near-infrared (NIR) video is a promising modality for contactless sleep monitoring, but recent…
Beyond the Verdict: Evidence-Aligned Evaluation of Visual Prompt-Injection Guardrails
arXiv:2609.05535v1 Announce Type: cross Abstract: Verdict-only evaluation does not reveal whether a vision-language model (VLM) used the visual evidence…
Robots Influencing Humans to Reveal their Goals during Collaboration and Competition
arXiv:2609.05519v1 Announce Type: cross Abstract: We propose a unified strategy for fast goal inference in human-robot interaction. The core idea is to…
When Agent Governance Helps
arXiv:2609.05531v1 Announce Type: cross Abstract: No specification says how a governed autotelic AI agent organization, where agents pursue self-generated…
Diffusion models for eye-gaze trajectory generation using position and velocity representations
arXiv:2609.05522v1 Announce Type: cross Abstract: Eye-tracking data are expensive to collect, requiring specialized hardware and controlled laboratory…
Situation Awareness for Intelligent Data Distribution in Connected Vehicles
arXiv:2609.05521v1 Announce Type: cross Abstract: The limitations of on-board sensors and blind spots caused by occlusion cause the reduction of…
DART: A DAG-Based Reputation and Incentive Framework via Blockchain-Enabled Governance for Trustworthy LLM Multi-Agent Collaboration
arXiv:2609.05529v1 Announce Type: cross Abstract: Large language model (LLM)-based multi-agent systems (MAS) predominantly rely on centralized…
ProToMEx: Rapid, Interpretable Explanations via Structured Representations
arXiv:2609.04265v1 Announce Type: cross Abstract: Existing post-hoc explainers for machine learning classifiers primarily focus on feature attribution,…
Emergent Goal-Directed Attention in Large Vision-Language Models
arXiv:2609.05517v1 Announce Type: cross Abstract: Human observers prioritize visual information according to task goals. Most computational models of…
