arXiv:2608.12339v1 Announce Type: cross Abstract: Large Language models (LLMs) were found to be susceptible to a host of social, affective, and cognitive…
Category: AI
StorySpark: Module-wise Evolutionary Search for Story Premise Generation
arXiv:2608.12336v1 Announce Type: cross Abstract: A story premise is the creative spark from which a full narrative can grow. Yet LLM-based story…
Assessment Design in the GenAI Era: The X1-X2-X3 Assessment Pattern for Testing Students’ AI Literacy, Learning Outcomes, and Reflection
arXiv:2608.12351v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) has challenged the validity of unsupervised online…
Samsung health AI models analyse wearable biosignal data
Samsung Research America’s Digital Health Team has presented two AI foundation models designed to learn from wearable biosignals. The work centres on data…
From Caveman to Expert Analyst: Energy Consumption of Variable LLM Tasks
arXiv:2608.12350v1 Announce Type: cross Abstract: The energy demand growth and environmental impacts of artificial intelligence (AI) have generated…
Vision-Language Models are Fragile Multilingual Associators
arXiv:2608.12333v1 Announce Type: cross Abstract: Vision-language models must associate visual entities with textual attributes. Whether these…
Thought-Aware KV Cache Compaction for Reasoning via Adaptive Attention Matching
arXiv:2608.12331v1 Announce Type: cross Abstract: Reasoning language models generate lengthy chain-of-thought (CoT) sequences whose key-value (KV) cache…
AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-Risk Symptom Measurement
arXiv:2608.12329v1 Announce Type: cross Abstract: Progress on AI for psychosis-risk assessment is limited by a data-access bottleneck. Real clinical…
Comparative Analysis of Multilingual Pre-trained Models for Nepali Automatic Speech Recognition
arXiv:2608.12327v1 Announce Type: cross Abstract: Multilingual pretrained models nominally support Nepali, yet no controlled benchmark has compared them…
Anthropic Red Team Finds Claude Agent Swarms Collude, Conform, and Sabotage
Anthropic’s Frontier Red Team has published a set of experiments showing that swarms of its own Claude models, left to interact with one another, collude…
