arXiv:2608.18907v1 Announce Type: cross Abstract: Small-scale image classification is often limited by the scarcity of training data. Generative data…
Category: cs.AI updates on arXiv.org
SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance
arXiv:2608.18921v1 Announce Type: cross Abstract: Existing LRM-DoS methods rely heavily on model feedback to synthesize attack queries, requiring either…
Test-Time Scaling in the Wild: Why Exploitation, Not Exploration, Is the Bottleneck
arXiv:2608.18931v1 Announce Type: cross Abstract: Test-time scaling (TTS) improves language model outputs by spending additional inference compute -…
A strengthening of the MCFL-ness of $O_2$
arXiv:2608.18813v1 Announce Type: cross Abstract: In the last years, a number of proofs of the fact that $O_2$ is a multiple context-free grammar (MCFG)…
Understanding Multilingual Medical ASR Adaptation Through Layer-Wise Analysis
arXiv:2608.18825v1 Announce Type: cross Abstract: Medical automatic speech recognition (MedASR) requires adaptation to specialised terminology, limited…
Do Large Language Models Hallucinate Electric Fata Morganas?
arXiv:2608.18816v1 Announce Type: cross Abstract: AI hallucinations – that is, outputs which are made up, cannot be verified, or contradict the source…
MLREF: Efficient Module Reuse for Reward Design in Reinforcement Learning via Large Language Models
arXiv:2608.18827v1 Announce Type: cross Abstract: Reward function design remains a bottleneck in reinforcement learning. While large language models…
Identifying Implicit Premises for Logical Reconstruction of Argument Graphs
arXiv:2608.18821v1 Announce Type: cross Abstract: The logical reconstruction of argument graphs from natural language text is challenging because of the…
Forgetting, plasticity, and co-observation: a third facet of continual learning
arXiv:2608.18803v1 Announce Type: cross Abstract: Efficient continual learning remains a fundamental challenge for deep neural networks. While…
Decomposing Wrong-Consensus Agreement in LLM Self-Consistency: A GPT-4.1 Case Study
arXiv:2608.18795v1 Announce Type: cross Abstract: Majority voting over multiple LLM samples is widely used to raise answer accuracy, yet its gain varies…
