arXiv:2608.15621v1 Announce Type: new Abstract: Human Activity Recognition (HAR) with self-administered wearables, such as at-home rehabilitation and…
Category: AI
Bias-Corrected Ceilings of Emotion Predictability from Human Label Variation Based on Instance-Level Fano Bounds
arXiv:2608.15619v1 Announce Type: new Abstract: Emotion recognition from text keeps improving on benchmarks, yet whether an accuracy ceiling has been…
VARM-Bench: Benchmarking Verifiable Structured Reasoning in Chinese Abusive Speech Moderation
arXiv:2608.15600v1 Announce Type: new Abstract: The widespread circulation of abusive online content has increased the need for reliable moderation of…
TRACE: Trajectory Aware Reasoning for Multi-Turn Adversarial Conversation Evaluation
arXiv:2608.15594v1 Announce Type: new Abstract: Multi-turn jailbreak attacks have emerged as a critical safety threat to LLMs, as harmful objectives are…
Argumentation for Common Ground: Finding Zones of Possible Agreement between Individuals in Conflict
arXiv:2608.15634v1 Announce Type: new Abstract: How can common ground between societies in conflict be identified when citizens’ acceptability of peace…
Admission Without Answers: Label-Free Certification and Experience Learning for LLM-Based Optimization Modeling
arXiv:2608.15565v1 Announce Type: new Abstract: Experience-learning agents for optimization modeling improve by storing verified skills, but existing…
ATLAS: Scaffold-Free Algorithm Synthesis by LLMs via Embedding-Guided Quality-Diversity Search
arXiv:2608.15546v1 Announce Type: new Abstract: Most LLM-based automated algorithm design methods optimize a designated component within a human-specified…
Agent Gym: A Framework for Continuous Evaluation and Evolution of LLM Agents Through Human-in-the-Loop Feedback
arXiv:2608.15591v1 Announce Type: new Abstract: Large Language Model (LLM) agents deployed in production environments face a fundamental tension: the…
Edcafe AI Review: The Teacher Tool That’ll Save Your Sunday
Ask any teacher what actually eats their time, and it’s rarely the hours in front of students. It’s everything that happens before and after: building a…
When Entropy Is Not Enough: Reclaiming Lost Semantics in LLM Output Length Prediction
arXiv:2608.15592v1 Announce Type: new Abstract: Efficient LLM serving is often bottlenecked by the need to pad sequences to a fixed maximum length, and…
