arXiv:2609.00818v1 Announce Type: new Abstract: We argue that financial report generation should operate at the analytical rather than structural level,…
Tag: cs.AI updates on arXiv.org
Verifiable Disaster Storylines and Causal Knowledge Graphs: A Citation-Grounded Pipeline from Heterogeneous Humanitarian Sources
arXiv:2609.00858v1 Announce Type: new Abstract: Effective humanitarian response depends on the rapid synthesis of heterogeneous, high-volume information…
Towards Generalizable Visually Grounded Exploration of Household Devices
arXiv:2609.00845v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have demonstrated impressive capabilities in static…
Polished but Unresolved: Identifying Late-Stage Pressure States in Long-Horizon Tool-Use Agents
arXiv:2609.00823v1 Announce Type: new Abstract: Long-horizon tool-use agents need not only to search and plan, but also to decide when to finalize. We…
When Features Become Instances: Inverted Contrastive Learning for Unsupervised Feature Selection
arXiv:2609.00782v1 Announce Type: new Abstract: Unsupervised feature selection seeks a compact subset of informative features without access to class…
Towards a Reliable and Practical Eval Pipeline
arXiv:2609.00805v1 Announce Type: new Abstract: LLM-based software systems increasingly require effective “evals” as quality gates in the development…
StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability?
arXiv:2609.00787v1 Announce Type: new Abstract: Humans need to study only a handful of well-written textbooks to master a discipline and attempt its…
One Policy, Any Budget: Internalizing Budget-Aware Search via Reinforcement Learning
arXiv:2609.00813v1 Announce Type: new Abstract: While reinforcement learning has enabled LLM-based search agents to invoke external tools, existing…
DiagEvo: Diagnosis-Guided Self-Evolution via Hierarchical Error Memory
arXiv:2609.00768v1 Announce Type: new Abstract: Self-play is an effective paradigm for language-model self-evolution, but without guidance, solver…
Escaping Redundant Reasoning: Structure-Aware Search for Inference-Time LLMs
arXiv:2609.00738v1 Announce Type: new Abstract: Inference-time search with large language models (LLMs) often concentrates on a small set of structurally…
