arXiv:2609.10950v1 Announce Type: cross Abstract: Recent multimodal sentiment analysis studies increasingly adopt text-centric fusion approaches to…
Author: script
EGGROLL, Unrolled: Understanding and Improving Low-Rank Evolution Strategies at Scale
arXiv:2609.10980v1 Announce Type: cross Abstract: EGGROLL makes evolution strategies (ES) practical for LLMs by replacing dense Gaussian weight…
A Mathematical Theory of Pragmatic Information
arXiv:2609.10986v1 Announce Type: cross Abstract: We propose a pragmatic information theory unifying communication, control, and decision-making. Its core…
AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds
Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model’s internal states, especially in the middle layers.…
Importance Weighting for Unlabeled-unlabeled Learning under Distribution Shift
arXiv:2609.10994v1 Announce Type: cross Abstract: Unlabeled-unlabeled (UU) learning allows us to learn a binary classifier from two sets of unlabeled data…
AI News Brief Hourly Summary 2026-09-12 16h : 12 posts
12 posts published in the last hour 13:32Evaluating Scaffolding-Oriented Multi-Agent Large Language Model System for Clinical Interview Training 13:32ReactHuman: A Physics-Grounded Benchmark for Human-Like Reactive Decision-Making in Embodied Multimodal LLMs 13:32DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt…
Evaluating Scaffolding-Oriented Multi-Agent Large Language Model System for Clinical Interview Training
arXiv:2609.10939v1 Announce Type: cross Abstract: Clinical education must prepare medical students to conduct safe and coherent patient interviews under…
ReactHuman: A Physics-Grounded Benchmark for Human-Like Reactive Decision-Making in Embodied Multimodal LLMs
arXiv:2609.10895v1 Announce Type: cross Abstract: Reacting to sudden physical hazards (catching a slipping plate, dodging a falling knife) is both a…
DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt Injection in LLM Agents
arXiv:2609.10892v1 Announce Type: cross Abstract: When an indirect prompt injection succeeds against an LLM agent, the compromise is visible in the…
AUC Maximization from Biased Positive-unlabeled Data with Confidence
arXiv:2609.10928v1 Announce Type: cross Abstract: Maximizing the area under the receiver operating characteristic curve (AUC) is a standard approach to…
