13 posts were published in the last hour
- 13:33 : StreamReason-Bench: Can Large Language Models Reason about Event-Time Stream-Processing Semantics?
- 13:33 : Mimicry without understanding: the origins of decision bias in large language models
- 13:33 : StorySpark: Module-wise Evolutionary Search for Story Premise Generation
- 13:33 : Assessment Design in the GenAI Era: The X1-X2-X3 Assessment Pattern for Testing Students’ AI Literacy, Learning Outcomes, and Reflection
- 13:33 : Samsung health AI models analyse wearable biosignal data
- 13:33 : From Caveman to Expert Analyst: Energy Consumption of Variable LLM Tasks
- 13:4 : Vision-Language Models are Fragile Multilingual Associators
- 13:4 : Thought-Aware KV Cache Compaction for Reasoning via Adaptive Attention Matching
- 13:4 : AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-Risk Symptom Measurement
- 13:4 : Comparative Analysis of Multilingual Pre-trained Models for Nepali Automatic Speech Recognition
- 13:4 : Anthropic Red Team Finds Claude Agent Swarms Collude, Conform, and Sabotage
- 13:4 : Steering the Language Axis: From Linear Decodability to Causal Control
- 13:0 : AI News Brief Hourly Summary 2026-08-14 15h : 16 posts