arXiv:2609.20816v1 Announce Type: cross Abstract: Professional design requires any-color control: the ability to specify an object’s target color with any…
Category: cs.AI updates on arXiv.org
Workspace Models: Lightweight Robotic Memory via Saliency-Driven Supervision
arXiv:2609.20820v1 Announce Type: cross Abstract: Complex robotic manipulation tasks frequently require a long-term memory of past events and actions. As…
FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations
arXiv:2609.20817v1 Announce Type: cross Abstract: Modeling articulated objects from sparse monocular views is challenging because each observation reveals…
Coding Agents with an Obstacle-Aware Harness for Safe Robot Manipulation
arXiv:2609.20822v1 Announce Type: cross Abstract: Coding agents have emerged as a promising paradigm for robot manipulation: a language model writes the…
Harm Laundering in GPT Models: Evidence That Gender Discrimination Is Transformed Rather Than Reduced Across Safety-Trained Generations
arXiv:2609.20779v1 Announce Type: cross Abstract: Safety evaluations for large language models rely on surface-form classifiers that report declining harm…
Semantic Action Graph: A Shared Representation for Agent Grounding and Human Interpretation of Sports Highlights
arXiv:2609.20768v1 Announce Type: cross Abstract: Generative agents are increasingly used to select and narrate video highlights, but they typically…
Quantifying Overclaiming Propensity in Frontier LLM Agents
arXiv:2609.20812v1 Announce Type: cross Abstract: Frontier coding agents are increasingly trusted to work autonomously for long periods, yet an agent’s…
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning
arXiv:2609.20784v1 Announce Type: cross Abstract: Multi-turn agents trained with reinforcement learning (RL) receive a single scalar reward per…
GeoAAC: Geometry-Based Adaptive Action Chunking from Denoising Trajectories in VLA Policies
arXiv:2609.20776v1 Announce Type: cross Abstract: Action chunking is widely used for action generation and execution in Vision-Language-Action (VLA)…
Prediction-Powered Smoothing and Validation for Disaggregated AI Evaluation
arXiv:2609.20758v1 Announce Type: cross Abstract: Evaluating an AI system requires disaggregated assessment, as performance varies across domains such as…
