arXiv:2609.09657v1 Announce Type: new Abstract: Existing emotional support conversation systems mainly focus on one-on-one seeker-supporter interactions…
Author: script
Black-Box Red Teaming of Agentic AI: A Taxonomy-Driven Framework for Automated Risk Discovery
arXiv:2609.09647v1 Announce Type: new Abstract: Agentic systems are rapidly moving to production, where they read untrusted inputs, call tools with real…
ContractEval: Query-Conditioned Execution Matching for Procedural Instruction Conformance
arXiv:2609.09458v1 Announce Type: new Abstract: As LLM agents move from answering questions to carrying out procedures, failures can be unwarranted rather…
CityPlanner: A Sandbox Agent for Executable Urban Planning
arXiv:2609.09578v1 Announce Type: new Abstract: Urban planning is a real-world spatial optimization problem that requires selecting feasible actions from…
From State Synchronization to Cognitive Self-Evolution: An Operational Architecture for Cognitive Digital Twins
arXiv:2609.09625v1 Announce Type: new Abstract: As Digital Twin (DT) systems evolve beyond state synchronization toward task-oriented and knowledge-driven…
Multi-Agent Agentic Graph Learning via Structural Signatures
arXiv:2609.09565v1 Announce Type: new Abstract: Agentic graph learning (AGL) has recently achieved promising results on graph reasoning tasks, where an…
A Function-Space Approach to the Statistical Mechanics of Learning Dynamics
arXiv:2609.09589v1 Announce Type: new Abstract: Deep neural networks exhibit regular macroscopic behavior despite highly nonlinear dynamics in vast…
AI News Brief Hourly Summary 2026-09-11 07h : 10 posts
10 posts published in the last hour 04:32Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations 04:32Valerant: An Automatic Navigable Game Map Generator via Action-Conditioned World Model Exploration 04:32XAI-Arena: Can LLMs Assess the Quality of XAI Explanations?…
Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations
arXiv:2609.09448v1 Announce Type: new Abstract: As agentic systems getting adopted rapidly in safety critical applications, it is vital to measure the…
Valerant: An Automatic Navigable Game Map Generator via Action-Conditioned World Model Exploration
arXiv:2609.09418v1 Announce Type: new Abstract: World Action Models (WAMs) couple predictive world modeling with action generation, allowing anticipated…
