arXiv:2608.12329v1 Announce Type: cross Abstract: Progress on AI for psychosis-risk assessment is limited by a data-access bottleneck. Real clinical…
Author: script
Comparative Analysis of Multilingual Pre-trained Models for Nepali Automatic Speech Recognition
arXiv:2608.12327v1 Announce Type: cross Abstract: Multilingual pretrained models nominally support Nepali, yet no controlled benchmark has compared them…
Anthropic Red Team Finds Claude Agent Swarms Collude, Conform, and Sabotage
Anthropic’s Frontier Red Team has published a set of experiments showing that swarms of its own Claude models, left to interact with one another, collude…
Steering the Language Axis: From Linear Decodability to Causal Control
arXiv:2608.12334v1 Announce Type: cross Abstract: Despite the impressive multilingual capabilities of Large Language Models, the latent dynamics dictating…
AI News Brief Hourly Summary 2026-08-14 15h : 16 posts
16 posts were published in the last hour 12:33 : LLMs Know the Constraint But Do Not Use It: Activation Bottlenecks in Pragmatic Constraint Reasoning 12:33 : When AI Is Right and the Process Is Wrong 12:33 : When AI…
LLMs Know the Constraint But Do Not Use It: Activation Bottlenecks in Pragmatic Constraint Reasoning
arXiv:2608.12321v1 Announce Type: cross Abstract: When a salient surface cue competes with an implicit feasibility constraint, LLMs often fail — but…
When AI Is Right and the Process Is Wrong
Imagine a common scenario in financial services. A team deploys AI to review contracts: hundreds of pages, repetitive clauses, and routine work that…
When AI Is Your Pastor: A Benchmark for Theological Triage and Pastoral Guidance in Large Language Models
arXiv:2608.12324v1 Announce Type: cross Abstract: People increasingly ask large language models (LLMs) for counsel on questions of faith, doctrine, and…
Your KV Cache Doesn’t Have a Bit Problem. It Has a Geometry Problem.
At identical 2-bit precision, one decision about which axis you quantize along swings a benchmark score from 2.88 to 63.53. Keys and values need opposite…
What Drives LLM Self-Reflection? A Controlled Ablation of Uncertainty Routing in Armed Conflict Forecasting
arXiv:2608.12322v1 Announce Type: cross Abstract: Self-reflection is widely assumed to improve LLM reasoning, yet which component drives the gain remains…
