arXiv:2609.24967v1 Announce Type: new Abstract: LLM agents are increasingly deployed in collaborative settings, yet long-term interaction may give rise to…
Category: cs.AI updates on arXiv.org
A Global Comparison of Schemas, Transparency, and Interoperability in Public-Sector AI Registers and Inventories
arXiv:2609.24883v1 Announce Type: new Abstract: Artificial intelligence (AI) registers and inventories aim to make governmental AI visible, but their…
BackTrend: Evaluating Scientific Weak-Signal Prediction via Backward Reconstruction
arXiv:2609.24921v1 Announce Type: new Abstract: Scientific weak signals are early, low-visibility research directions that later become central to mature…
Pinocchio: Fast Uncertainty Estimates for Black-Box Language Models
arXiv:2609.24881v2 Announce Type: new Abstract: In high-stakes decision-making applications of large language models (LLMs), practitioners require not…
Et Tu, Brute? Economic Misalignment in Personal AI Agents
arXiv:2609.24927v1 Announce Type: new Abstract: Personal AI agents make recommendations and take actions on people’s behalf in high-stakes economic…
Partner-Specific Affective Precision in Social Active Inference
arXiv:2609.24876v1 Announce Type: new Abstract: In multi-agent social settings, model reliability varies across relationships. Beyond inferring what…
Convex AI Compositionality and the Governance of AI System Populations
arXiv:2609.24784v1 Announce Type: new Abstract: AI governance increasingly requires providers and public authorities to reason about multiple AI…
GRUET: Quantifying Uncertainty of Agentic Reasoning-and-Acting Processes
arXiv:2609.24831v1 Announce Type: new Abstract: Agents have attracted considerably increasing attention due to the power of executing both Reasoning and…
MedRSI: Recursive Self-Improvement for Medical Agents via Clinically Aligned Self-Evolution
arXiv:2609.24838v1 Announce Type: new Abstract: Medical agents increasingly combine general reasoning models with specialized clinical tools, yet their…
Extracting Arguments, Not Just Classifying Them: Instruction-Tuned LLMs for Generative Component Detection
arXiv:2609.24855v1 Announce Type: new Abstract: Argumentative component detection (ACD) is a core subtask of Argument(ation) Mining (AM) and one of its…
