arXiv:2609.04197v1 Announce Type: cross Abstract: Evolutionary prompt optimizers such as GEPA suffer from prompt bloat: each iteration appends rules and…
Tag: cs.AI updates on arXiv.org
Compile by Training: Turning Natural-Language Specifications into Local Neural Functions
arXiv:2609.04199v1 Announce Type: cross Abstract: Many recurring text functions are easy to describe but difficult to implement with rules, while calling…
Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views
arXiv:2609.04180v1 Announce Type: cross Abstract: Gaps remain in our understanding of how large language models (LLMs) acquire knowledge during…
One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing
arXiv:2609.04190v1 Announce Type: cross Abstract: Video editing spans diverse editing paradigms, yet achieving high-quality instruction-guided and…
SWE-Gate: Passing Functional Tests Is Not Enough for Software Engineering Agents
arXiv:2609.04167v1 Announce Type: cross Abstract: Repository-level software engineering benchmarks have significantly advanced the evaluation of coding…
Adaptive Vision-Language Grasping via Composable Foundation Priors and Generalizable Grasp Synthesis
arXiv:2609.04096v1 Announce Type: cross Abstract: This paper proposes AdaRoboVLG, a task-adaptive Vision-Language-Grasp (VLG) framework that supports…
A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle
arXiv:2609.04147v1 Announce Type: cross Abstract: This paper presents a low-cost, open experimental platform for research in end-to-end autonomous driving…
SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center
arXiv:2609.04159v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly proposed as autonomous SOC analysts, but two…
Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR
arXiv:2609.04108v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and on-policy distillation (OPD) have emerged as…
CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation
arXiv:2609.04083v1 Announce Type: cross Abstract: MLLM-based embedding models remain limited in compositional retrieval, often failing to distinguish…
