arXiv:2609.05245v1 Announce Type: new Abstract: Human knowledge is inherently structured and interdependent: mastery of a concept requires prior mastery…
Substrate-Aware AI Agents: Execution Context as a First-Class Input
arXiv:2609.05232v1 Announce Type: new Abstract: Autonomous AI agents increasingly select actions in environments whose memory, execution-time, runtime,…
The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior
arXiv:2609.05190v1 Announce Type: new Abstract: In this paper we illustrate a novel architecture generating interpretable behavior and explanations. We…
CABAL: Multi-Agent Simulacra for Tracing the Effects of Collusive Bidding in Peer Review
arXiv:2609.05227v1 Announce Type: new Abstract: Recent reports during the AAAI-27 review cycle highlight the risk of reviewers coordinating bids for…
ACE: Adaptive Calibration-Free Expert Skipping for MoE-based LLMs
arXiv:2609.05228v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures provide an efficient paradigm for scaling large language models…
What Matters in On-Policy Distillation? A Perspective on Data Efficiency and Data Selection
arXiv:2609.05198v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has emerged as a widely adopted post-training paradigm for enhancing large…
AI News Brief Hourly Summary 2026-09-07 14h : 10 posts
10 posts published in the last hour 11:33SciDocBench: A Workflow-Centered Benchmark and Data Pipeline for Scientific Document Understanding 11:33A Hybrid Predictive Ensemble of Machine Learning and Deep Neural Networks for Early Cardiovascular Disease Risk Assessment 11:33ProCA: Progressive Contrastive Alignment for…
SciDocBench: A Workflow-Centered Benchmark and Data Pipeline for Scientific Document Understanding
arXiv:2609.05141v1 Announce Type: new Abstract: Scientific papers require models to reason jointly over text, equations, figures, tables, code, and…
A Hybrid Predictive Ensemble of Machine Learning and Deep Neural Networks for Early Cardiovascular Disease Risk Assessment
arXiv:2609.05146v1 Announce Type: new Abstract: This study introduces an intelligent framework that integrates machine learning and deep neural network…
ProCA: Progressive Contrastive Alignment for Robust EEG Visual Decoding
arXiv:2609.05094v1 Announce Type: new Abstract: Electroencephalogram (EEG) visual decoding aims to recover visual semantics from non-invasive neural…
Compact Bellman-Grounded Cognitive Maps for Cost-Aware Navigation
arXiv:2609.05104v1 Announce Type: new Abstract: Biological agents navigate familiar environments not by re-solving routes for each new goal, but by…
Unifying ICL, SFT, KL-Regularized RL Through a Bayesian Lens
arXiv:2609.05111v1 Announce Type: new Abstract: Large language models are now trained and evaluated under a diverse set of paradigms: supervised…
TruthInsightBench: An Evidence-Grounded Benchmark for Automated Evaluation of Open-Ended Scientific Discovery Agents
arXiv:2609.05079v1 Announce Type: new Abstract: Autonomous coding agents are increasingly proposed as AI-scientist systems that conduct analyses and write…
Constructing and Evaluating Clinical Reasoning Trajectories for Medical Agent
arXiv:2609.05090v1 Announce Type: new Abstract: Evaluation of medical artificial intelligence agents remains predominantly answer-centric, assessing only…
LLM-Guided Program Evolution for Circle Packing: Breaking 10 Packomania Records for $28
arXiv:2609.05093v1 Announce Type: new Abstract: We present Discovery Loop, a lightweight system that uses a large language model (LLM) to iteratively…
Measuring AI Accountability Through Argumentation Analysis: Can Model Reasoning Withstand Scrutiny?
arXiv:2609.05088v1 Announce Type: new Abstract: AI oversight methods rely on ground truth for validation, but what constitutes appropriate AI behavior is…
MePo++: Unifying Representation Refinement and Reconciliation for General Continual Learning
arXiv:2609.05075v1 Announce Type: new Abstract: General continual learning (GCL) aims to learn from evolving data streams without task identities,…
AI News Brief Hourly Summary 2026-09-07 13h : 13 posts
13 posts published in the last hour 10:33Language models judge war differently when tested for alignment 10:33Towards Efficient Evaluation of Evolutionary Transfer Optimization: Case Studies on Task-Parameterized Applications 10:33Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for…
