arXiv:2608.12336v1 Announce Type: cross Abstract: A story premise is the creative spark from which a full narrative can grow. Yet LLM-based story…
Assessment Design in the GenAI Era: The X1-X2-X3 Assessment Pattern for Testing Students’ AI Literacy, Learning Outcomes, and Reflection
arXiv:2608.12351v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) has challenged the validity of unsupervised online…
Samsung health AI models analyse wearable biosignal data
Samsung Research America’s Digital Health Team has presented two AI foundation models designed to learn from wearable biosignals. The work centres on data…
From Caveman to Expert Analyst: Energy Consumption of Variable LLM Tasks
arXiv:2608.12350v1 Announce Type: cross Abstract: The energy demand growth and environmental impacts of artificial intelligence (AI) have generated…
Vision-Language Models are Fragile Multilingual Associators
arXiv:2608.12333v1 Announce Type: cross Abstract: Vision-language models must associate visual entities with textual attributes. Whether these…
Thought-Aware KV Cache Compaction for Reasoning via Adaptive Attention Matching
arXiv:2608.12331v1 Announce Type: cross Abstract: Reasoning language models generate lengthy chain-of-thought (CoT) sequences whose key-value (KV) cache…
AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-Risk Symptom Measurement
arXiv:2608.12329v1 Announce Type: cross Abstract: Progress on AI for psychosis-risk assessment is limited by a data-access bottleneck. Real clinical…
Comparative Analysis of Multilingual Pre-trained Models for Nepali Automatic Speech Recognition
arXiv:2608.12327v1 Announce Type: cross Abstract: Multilingual pretrained models nominally support Nepali, yet no controlled benchmark has compared them…
Anthropic Red Team Finds Claude Agent Swarms Collude, Conform, and Sabotage
Anthropic’s Frontier Red Team has published a set of experiments showing that swarms of its own Claude models, left to interact with one another, collude…
Steering the Language Axis: From Linear Decodability to Causal Control
arXiv:2608.12334v1 Announce Type: cross Abstract: Despite the impressive multilingual capabilities of Large Language Models, the latent dynamics dictating…
AI News Brief Hourly Summary 2026-08-14 15h : 16 posts
16 posts were published in the last hour 12:33 : LLMs Know the Constraint But Do Not Use It: Activation Bottlenecks in Pragmatic Constraint Reasoning 12:33 : When AI Is Right and the Process Is Wrong 12:33 : When AI…
LLMs Know the Constraint But Do Not Use It: Activation Bottlenecks in Pragmatic Constraint Reasoning
arXiv:2608.12321v1 Announce Type: cross Abstract: When a salient surface cue competes with an implicit feasibility constraint, LLMs often fail — but…
When AI Is Right and the Process Is Wrong
Imagine a common scenario in financial services. A team deploys AI to review contracts: hundreds of pages, repetitive clauses, and routine work that…
When AI Is Your Pastor: A Benchmark for Theological Triage and Pastoral Guidance in Large Language Models
arXiv:2608.12324v1 Announce Type: cross Abstract: People increasingly ask large language models (LLMs) for counsel on questions of faith, doctrine, and…
Your KV Cache Doesn’t Have a Bit Problem. It Has a Geometry Problem.
At identical 2-bit precision, one decision about which axis you quantize along swings a benchmark score from 2.88 to 63.53. Keys and values need opposite…
What Drives LLM Self-Reflection? A Controlled Ablation of Uncertainty Routing in Armed Conflict Forecasting
arXiv:2608.12322v1 Announce Type: cross Abstract: Self-reflection is widely assumed to improve LLM reasoning, yet which component drives the gain remains…
The Shift from AI Capability to AI Control
A series of high-profile incidents involving frontier AI providers has left businesses asking a question that, until recently, many hadn’t seriously…
The AI Accountability Ecosystem in the Era of Language Models
arXiv:2608.12320v1 Announce Type: cross Abstract: This article reviews and updates the framework for accountability in AI based on account- ability…
