arXiv:2608.30226v1 Announce Type: new Abstract: Modular compression has enabled considerable parameter reduction in LLMs while preserving strong language…
Tag: cs.AI updates on arXiv.org
VERA: Authority-Preserving Edge Revocation for Federated AI-Agent Workflows
arXiv:2608.30091v1 Announce Type: new Abstract: Modern agent frameworks compose planners, tool agents, remote services, and shared specialists into…
SPARK: Skeleton-Guided Reasoning Synthesis from Large-Scale Scientific Literature
arXiv:2608.30214v1 Announce Type: new Abstract: Scientific reasoning remains challenging for open-source models, largely due to the lack of high-quality…
FaVOR: LLM-Based Agentic Framework for Factor Mining via Empirical Validation
arXiv:2608.30192v1 Announce Type: new Abstract: Traditional finance relies on experts to hand-craft factors through a principled process grounded in…
Game-Agnostic Value Functions through Automatic JSON Feature Extraction
arXiv:2608.30056v1 Announce Type: new Abstract: JSON Bag-of-Tokens (JSON-Bag) is a recently proposed method to generically represent game trajectories by…
Spec2Twin-Chain: Orchestrating Bi-Level Optimization with LLMs for Blockchain Digital Twin Construction
arXiv:2608.30050v1 Announce Type: new Abstract: Building a blockchain digital twin largely requires translating domain knowledge and specific system…
Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide
arXiv:2608.30051v1 Announce Type: new Abstract: Process reward models (PRMs) provide dense step-level guidance for search-based reasoning, enabling…
Can LLM Agents Discover? Evaluating Creativity on ML Engineering Tasks
arXiv:2608.30047v1 Announce Type: new Abstract: Recent AI systems promise autonomous scientific discovery, claiming to discover algorithms and produce…
Balance of Benchmarks: Semantic Density Reweighting for Benchmark Multiplicity and Task-Conditioned Evaluation
arXiv:2608.30044v1 Announce Type: new Abstract: Language models are commonly compared by averaging scores across a benchmark list with equal weight. Such…
Beyond Uncertainty: Multi-Solver Disagreement Rewards for Self-Evolving Reasoning Curricula
arXiv:2608.30035v1 Announce Type: new Abstract: Self-evolving reasoning frameworks train a Challenger to generate questions exposing a Solver’s…
