arXiv:2608.12320v1 Announce Type: cross Abstract: This article reviews and updates the framework for accountability in AI based on account- ability…
Tag: cs.AI updates on arXiv.org
Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance
arXiv:2608.12323v1 Announce Type: cross Abstract: Specifying a penalty can paradoxically convert a legal obligation into a cost-benefit calculation that…
QuoteBench: How Matched Scores Can Hide Command-Path Failures
arXiv:2608.13547v1 Announce Type: new Abstract: LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model…
AlayaWorld: Interactive Long-Horizon World Modeling – Full Technical Report (v1.1)
arXiv:2608.13492v1 Announce Type: new Abstract: This report presents an improved version of AlayaWorld. While the backbone architecture, chunk-wise…
A Unifying Perspective on Causal World Models: From Observations to Representations to Structure
arXiv:2608.13456v1 Announce Type: new Abstract: World Models (WM) are increasingly seen as a foundation for intelligent agents that can predict, plan, and…
OmniScientist: An Omni-Modal Omni-Discipline AI Scientist
arXiv:2608.13558v1 Announce Type: new Abstract: Recent advances in foundation models have enabled AI scientists to automate increasingly complete research…
MARC v1: An Open-Source Multi-Agent Framework for Clinical AI Reasoning and Coordination
arXiv:2608.13476v1 Announce Type: new Abstract: We present Multi-Agent Reasoning and Coordination (MARC), an open-source framework that replaces…
Enhancing Virtual Agents through SLMs and Edge-Computing: An Exploratory Evaluation of Think and Memory Processes
arXiv:2608.13420v1 Announce Type: new Abstract: Embodied intelligent virtual agents are expected to operate as persistent, adaptive, and context-aware…
Beyond Final Scores: A Systematic Evaluation of Agents for Long-Horizon AI Research and Development
arXiv:2608.13417v1 Announce Type: new Abstract: Autonomous agents are increasingly capable of improving models, systems, and other technical artifacts…
RAIL: An Automatic Classifier of the Artificial Intelligence Readiness Level
arXiv:2608.13428v1 Announce Type: new Abstract: Assessing the maturity of artificial intelligence technologies is essential for investment decisions,…
