arXiv:2609.18708v1 Announce Type: cross Abstract: In reinforcement learning for large language models, Proximal Policy Optimization (PPO) commonly uses a…
Tag: cs.AI updates on arXiv.org
Using OCR Heads to Verbalize Image Semantics
arXiv:2609.18823v1 Announce Type: cross Abstract: How do VLMs map from pixels to semantics? To understand this general question, we focus on a narrow one:…
Echo: Learning-based Matching Decompilation using Trusted Back Translation
arXiv:2609.18706v1 Announce Type: cross Abstract: Neural decompilers can recover readable and recompilable source code from binaries, but their…
A Scalable Framework for Automated NER Annotation Correction in Low-Resource Languages
arXiv:2609.18739v1 Announce Type: cross Abstract: Poor quality or noisy annotations in Named Entity Recognition (NER), as in any other NLP task, make it…
ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks
arXiv:2609.18805v1 Announce Type: cross Abstract: Coding agents are typically evaluated with desired behavior specified through issues or instructions. In…
Beyond EER: Multi-Dimensional Evaluation of Information Leakage in Speaker De-Identification
arXiv:2609.18673v1 Announce Type: cross Abstract: Speaker de-identification (SDID) aims to preserve privacy by concealing speaker identity while…
CoRe-MARL: Cooperative Redistribution Under Unknown Dynamics Using Recurrent Multi-Agent Reinforcement Learning
arXiv:2609.18639v1 Announce Type: cross Abstract: Emergency management assistance programs, such as relief distribution, are essential for delivering…
GenStream: Semantic Streaming Framework for Generative Reconstruction of Human-centric Media
arXiv:2609.18634v1 Announce Type: cross Abstract: Video streaming dominates global internet traffic, yet conventional pipelines remain inefficient for…
Generalist-Specialist Mixture-of-Experts for Rare Pathology Detection in Multimodal Imaging
arXiv:2609.18688v1 Announce Type: cross Abstract: AI models for multimodal medical imaging must balance modality-specific specialization with cross-modal…
PACT: Can Enterprise AI Assistants Be Trusted Under Pressure?
arXiv:2609.18605v1 Announce Type: cross Abstract: As corporate AI adoption continues to grow, enterprise-grade LLM agents are being deployed into…
