arXiv:2609.01409v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance in generating scientific figures from text or…
Category: cs.AI updates on arXiv.org
Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers
arXiv:2609.01466v1 Announce Type: new Abstract: A long-horizon agent’s trace outgrows both of its consumers: the human observer monitoring the run, and…
Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement
arXiv:2609.01481v1 Announce Type: new Abstract: This paper studies autonomous software development, in which LLM-based coding agents transform high-level…
EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems
arXiv:2609.01360v1 Announce Type: new Abstract: Large language model (LLM) agent failures often contain multiple related errors rather than a single…
Automated Event Log Generation from Unstructured Text Using Finetuned LLMs
arXiv:2609.01320v1 Announce Type: new Abstract: Process mining (PM) provides a powerful framework for discovering and optimizing operational processes…
SymFold: Synergizing Evolutionary and Structural Priors for Accurate Protein Inverse Folding
arXiv:2609.01353v1 Announce Type: new Abstract: Protein inverse folding aims to recover amino acid sequences for a given 3D protein structure,…
LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting
arXiv:2609.01337v1 Announce Type: new Abstract: LLM-based forecasting systems have improved on real-world tasks such as financial markets and sports…
Cheap Verifiers, Large Blind Spots: Measuring the Reliability Cost of Cost-Saving Cascades
arXiv:2609.01345v1 Announce Type: new Abstract: Inference cascades cut cost by answering most queries with a cheap model and escalating a hard tail to a…
A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation
arXiv:2609.01315v1 Announce Type: new Abstract: Building an omni-modal foundation model means evaluating it across text, image, video, and audio.…
Making Prospective Memory SLM-Shaped: Typed Intention Stores for Small-Model Agents
arXiv:2609.01272v1 Announce Type: new Abstract: Prospective memory means carrying out a deferred intention at the right future cue while other work…
