arXiv:2608.24893v2 Announce Type: replace Abstract: In competitive first-person shooter (FPS) games such as Counter-Strike 2 (CS2), account-integrity…
Category: cs.AI updates on arXiv.org
Jiuge-Tuiqiao: An Interpretable Human-AI System for Classical Chinese Poetry Refinement
arXiv:2608.23098v2 Announce Type: replace Abstract: Classical Chinese poetry composition has long valued Tuiqiao, the iterative refinement of words,…
From Inertia to Objectivity: Improving Deep Research Agents with Noise Isolation
arXiv:2608.23045v2 Announce Type: replace Abstract: Web search agents powered by Large Language Models (LLMs) show strong promise, but deep research tasks…
FlavourBench: Executable Culinary Reward Maps for Language Model Evaluation and Post-Training
arXiv:2608.20574v2 Announce Type: replace Abstract: Open-ended language-model evaluation often substitutes another model or a small preference panel for a…
NiyamAI – An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs
arXiv:2608.07167v2 Announce Type: replace Abstract: Autonomous LLM agents with tool execution capabilities introduce severe security risks through prompt…
Blast Radius
arXiv:2608.07440v3 Announce Type: replace Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a…
SPAR-Hate: Auditor-Guided Multi-Perspective Role Reasoning for Bilingual Hate Speech Parsing
arXiv:2608.22018v3 Announce Type: replace Abstract: Hate speech research has moved from coarse-grained classification towards structured parsing, where…
Beyond Endpoint Gains: A Weight-Delta Audit of Medical Specialization
arXiv:2608.20768v3 Announce Type: replace Abstract: Specialist language models are usually understood through endpoint gains: the generalist scores lower,…
Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observations
arXiv:2608.03611v2 Announce Type: replace Abstract: Multimodal Sentiment Analysis (MSA) integrates text, audio, and vision to infer human affect, yet…
Heaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language Models
arXiv:2607.04562v2 Announce Type: replace Abstract: Large language models (LLMs) generate fluent outputs that can be wrong. Unlike humans, who often…
