arXiv:2609.18849v1 Announce Type: cross Abstract: An agentic request spends substantial wall-clock time waiting for tools, and its KV cache holds GPU…
Category: cs.AI updates on arXiv.org
Decodable but Misrouted: Sparse Features Uncover a Readout Gap in Vision-Language Models for Harmful Meme Detection
arXiv:2609.18860v1 Announce Type: cross Abstract: When a large vision-language model misclassifies a harmful meme, the failure may reflect missing…
ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions
arXiv:2609.18864v1 Announce Type: cross Abstract: Privacy evaluations of tool-using LLM agents often inspect a designated action, final response, or…
GrainSpeech: Less Context, More Detail for Compact Speech Synthesis
arXiv:2609.18856v1 Announce Type: cross Abstract: Compact acoustic models face a challenging quality-capacity trade-off. We investigate two factors in…
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening
arXiv:2609.18708v1 Announce Type: cross Abstract: In reinforcement learning for large language models, Proximal Policy Optimization (PPO) commonly uses a…
Using OCR Heads to Verbalize Image Semantics
arXiv:2609.18823v1 Announce Type: cross Abstract: How do VLMs map from pixels to semantics? To understand this general question, we focus on a narrow one:…
Echo: Learning-based Matching Decompilation using Trusted Back Translation
arXiv:2609.18706v1 Announce Type: cross Abstract: Neural decompilers can recover readable and recompilable source code from binaries, but their…
A Scalable Framework for Automated NER Annotation Correction in Low-Resource Languages
arXiv:2609.18739v1 Announce Type: cross Abstract: Poor quality or noisy annotations in Named Entity Recognition (NER), as in any other NLP task, make it…
ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks
arXiv:2609.18805v1 Announce Type: cross Abstract: Coding agents are typically evaluated with desired behavior specified through issues or instructions. In…
Beyond EER: Multi-Dimensional Evaluation of Information Leakage in Speaker De-Identification
arXiv:2609.18673v1 Announce Type: cross Abstract: Speaker de-identification (SDID) aims to preserve privacy by concealing speaker identity while…
