arXiv:2608.30912v1 Announce Type: new Abstract: Artificial intelligence (AI) and natural language processing (NLP) are increasingly used to extract,…
Tag: cs.AI updates on arXiv.org
CARVE: Verified Expansion for Variable-Length Generation in Diffusion Language Models
arXiv:2608.30922v1 Announce Type: new Abstract: Masked diffusion language models predict tokens from a partially observed response canvas, enabling…
CAER: Causal Action Effect Reweighting for World Model Training
arXiv:2608.30897v1 Announce Type: new Abstract: World models are becoming core infrastructure for embodied intelligence, with action-conditioned video…
VFR-Audit: Verdict-Level Reliability for Fairness Audits in Hospital Length-of-Stay Prediction
arXiv:2608.30846v1 Announce Type: new Abstract: Fairness audits in clinical Artificial Intelligence convert continuous fairness metrics into binary…
Which Rules Matter Now? Policy-Centroid Routing Before an Intelligent System Acts
arXiv:2608.30757v1 Announce Type: new Abstract: Before an intelligent system can decide whether an action is allowed, it must first know which rules the…
HSRM: Hidden-State Reward Models for Test-Time Verification
arXiv:2608.30841v1 Announce Type: new Abstract: Large language models can often generate plausible mathematical reasoning traces, but reliably identifying…
Autoregressive Mosaics: Probing 2D Spatial Reasoning in Text-Only Language Models
arXiv:2608.30751v1 Announce Type: new Abstract: Large language models (LLMs) trained only on text and code can sometimes generate programs that draw…
Multimodal Adaptive Expert Selection with Text Routing and Ordinal Prototype Optimization for Sentiment Analysis
arXiv:2608.30726v1 Announce Type: new Abstract: Multimodal Sentiment Analysis (MSA) is a fundamental component of affective computing that aims to…
SkillZip Pro: Execution-Aware Dynamic Compression of Progressively Loaded Skills for Self-Evolving Agents
arXiv:2608.30785v1 Announce Type: new Abstract: Production agent skills are directory bundles, not isolated prompts. The root is loaded at activation;…
MedAgent-R1: Faithfulness-Aware Reinforcement Learning for Evidence-Grounded Medical Reasoning
arXiv:2608.30676v1 Announce Type: new Abstract: When medical AI systems hallucinate clinical reasoning, the consequences extend beyond incorrect answers:…
