arXiv:2609.20625v1 Announce Type: cross Abstract: Large language model responses are non-deterministic, so failures in LLM agents are hard to reproduce: a…
Tag: cs.AI updates on arXiv.org
Large Language Models as Falsifiers for Cyber-Physical Systems
arXiv:2609.20752v1 Announce Type: cross Abstract: Falsification searches for counterexamples to formal specifications in cyber-physical systems (CPS).…
Don’t Mask the Environment: Observation Supervision Changes How Agents Explore Under RL
arXiv:2609.20715v1 Announce Type: cross Abstract: Agent trajectories record what an agent does and what happens next. Yet standard supervised fine-tuning…
HIL-UMI: Bringing Human-in-the-Loop Post-Training of Vision-Language-Action Models to Universal Manipulation Interface
arXiv:2609.20659v1 Announce Type: cross Abstract: Large-scale vision-language-action (VLA) models provide powerful priors for robot manipulation, yet…
Model-Agnostic and Language-Agnostic Voice Pipeline Improvement for the Agriculture Domain
arXiv:2609.20504v1 Announce Type: cross Abstract: FarmerChat is Digital Green’s AI-powered agricultural advisory assistant for smallholder farmers, who…
Accelerating Visual Policy Learning with Sampling-Based Model Predictive Control
arXiv:2609.20575v1 Announce Type: cross Abstract: Learning visual policies for locomotion and manipulation requires coordinating contact with the…
Inference-Engine Fingerprinting Attacks are Practical: Exploring Model-Driven Environmental Discovery, Exploitation, and Escape
arXiv:2609.20614v1 Announce Type: cross Abstract: Frontier AI models are rapidly gaining the ability to exploit vulnerabilities in complex pieces of…
A Simulation Platform for AUV Fault Recovery: Exploring LLM-Based Diagnostic Strategies
arXiv:2609.20620v1 Announce Type: cross Abstract: Autonomous underwater vehicles (AUVs) operating beyond reliable communications must recover from…
Mitigating Retaliatory Algorithmic Collusion in Repeated Games
arXiv:2609.20548v1 Announce Type: cross Abstract: Reinforcement learning agents trained to maximize their own reward in repeated interactions can converge…
Fingerprinting Multimodal Large Language Models
arXiv:2609.20457v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) enable a wide range of image-text reasoning tasks, recent…
