arXiv:2606.01444v2 Announce Type: replace Abstract: Scientific discovery is not only answer generation but revision of the representational regime in…
Category: cs.AI updates on arXiv.org
Weak Critics Make Strong Learners: On-Policy Critique Distillation for Scalable Oversight
arXiv:2606.00424v2 Announce Type: replace Abstract: As large language models become stronger, weak supervisors may fail to provide reliable labels,…
GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents
arXiv:2605.29668v3 Announce Type: replace Abstract: LLM agents acting in structured environments fail in operational rather than conversational ways, and…
SEISMO: Explanation-Aware, Trajectory-Conditioned LLM Agents for Sample-Efficient Molecular Optimisation
arXiv:2602.00663v3 Announce Type: replace Abstract: Optimizing molecules to achieve desired properties is a central bottleneck across the chemical…
Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning
arXiv:2504.09772v3 Announce Type: replace Abstract: Test-Time Scaling has emerged as a powerful method to extend the reasoning capabilities of Large…
ACQ: A Deployed Two-Stage Framework for Automated Creative Quota Allocation in Large-Scale Online Advertising
arXiv:2412.06167v2 Announce Type: replace Abstract: In digital advertising, demand-side platforms (DSPs) allow advertisers to create multiple ad creatives…
Recognizing Artificial Minds: A Philosophical Defense of AI Cognition
arXiv:2504.13988v2 Announce Type: replace Abstract: This work defends the ‘Whole Hog Thesis’: sophisticated Large Language Models (LLMs) like ChatGPT are…
Online design of dynamic networks
arXiv:2410.08875v3 Announce Type: replace Abstract: Designing a network (e.g., a telecommunication or transport network) is mainly done offline, in a…
Re$^3$Cap: Retrieval-Guided Refinement for Image Captioning Enhancement via Reinforcement Learning
arXiv:2608.21305v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has demonstrated significant gains in image captioning, yet it is still…
TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems
arXiv:2608.21343v1 Announce Type: cross Abstract: Contextualization is essential for production automatic speech recognition (ASR) systems, where…
