arXiv:2609.10986v1 Announce Type: cross Abstract: We propose a pragmatic information theory unifying communication, control, and decision-making. Its core…
AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds
Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model’s internal states, especially in the middle layers.…
Importance Weighting for Unlabeled-unlabeled Learning under Distribution Shift
arXiv:2609.10994v1 Announce Type: cross Abstract: Unlabeled-unlabeled (UU) learning allows us to learn a binary classifier from two sets of unlabeled data…
AI News Brief Hourly Summary 2026-09-12 16h : 12 posts
12 posts published in the last hour 13:32Evaluating Scaffolding-Oriented Multi-Agent Large Language Model System for Clinical Interview Training 13:32ReactHuman: A Physics-Grounded Benchmark for Human-Like Reactive Decision-Making in Embodied Multimodal LLMs 13:32DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt…
Evaluating Scaffolding-Oriented Multi-Agent Large Language Model System for Clinical Interview Training
arXiv:2609.10939v1 Announce Type: cross Abstract: Clinical education must prepare medical students to conduct safe and coherent patient interviews under…
ReactHuman: A Physics-Grounded Benchmark for Human-Like Reactive Decision-Making in Embodied Multimodal LLMs
arXiv:2609.10895v1 Announce Type: cross Abstract: Reacting to sudden physical hazards (catching a slipping plate, dodging a falling knife) is both a…
DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt Injection in LLM Agents
arXiv:2609.10892v1 Announce Type: cross Abstract: When an indirect prompt injection succeeds against an LLM agent, the compromise is visible in the…
AUC Maximization from Biased Positive-unlabeled Data with Confidence
arXiv:2609.10928v1 Announce Type: cross Abstract: Maximizing the area under the receiver operating characteristic curve (AUC) is a standard approach to…
GPT-6 Astra needs leaner prompts and fewer guardrails, OpenAI recommends
Overly long skill descriptions, blanket reading requirements, and rigid approval rules can get in GPT-6 Astra’s way, warns OpenAI’s Eric Provencher. More…
Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current Architectures
arXiv:2609.10893v1 Announce Type: cross Abstract: Recent advances in large language models have transformed human-computer interaction. Despite their…
No-Box Vulnerability Analysis: Description-only Detection of Indirect Prompt Injection Vulnerabilities in MCP Servers
arXiv:2609.10854v1 Announce Type: cross Abstract: Conventional vulnerability analysis relies on either system access or dynamic interaction, all of which…
Counterfactual Marginalisation: Framework for Evaluating Robustness to Nuisance Variables
arXiv:2609.10778v1 Announce Type: cross Abstract: Machine learning models can achieve strong test performance while relying on demographic or…
Story Imprinting: AI Assistants Absorb Traits from Human Characters They Resemble
arXiv:2609.10883v1 Announce Type: cross Abstract: Language models are trained to implement a helpful AI Assistant character (e.g., Claude). We explore how…
Are We Really Doing Few-Shot Learning? A Critical Examination of Pre-Training Assumptions
arXiv:2609.10851v1 Announce Type: cross Abstract: Few-shot learning is commonly evaluated under protocols that pre-train a model on a large auxiliary set…
Tapes Together Strong: The Co-evolution of Computation and Cooperation
arXiv:2609.10817v1 Announce Type: cross Abstract: How does cooperation evolve in complex agentic systems? Prior work in evolutionary game theory studies…
AI News Brief Hourly Summary 2026-09-12 15h : 11 posts
11 posts published in the last hour 12:32Temporal and Multimodal Deep Learning for Cyberattack Detection in LEO Satellite Systems 12:32When Synthetic Data Hurts: On Catastrophic Forgetting in Skill Retrieval for LLM Agents 12:32Beyond Static Guarantees: Measuring the Static-Pass Dynamic-Fail Gap…
Temporal and Multimodal Deep Learning for Cyberattack Detection in LEO Satellite Systems
arXiv:2609.10746v1 Announce Type: cross Abstract: The growing reliance on Low-Earth Orbit (LEO) satellite communication systems has increased the need for…
When Synthetic Data Hurts: On Catastrophic Forgetting in Skill Retrieval for LLM Agents
arXiv:2609.10750v1 Announce Type: cross Abstract: LLM agents increasingly rely on external skills retrieved at runtime, making skill selection from large…
