arXiv:2609.13396v1 Announce Type: new Abstract: Multi-objective Bayesian optimisation (MOBO) is a sample-efficient approach for optimising expensive…
Category: cs.AI updates on arXiv.org
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
arXiv:2609.13356v1 Announce Type: new Abstract: In this work, we present ZGCM-1, a fully open 7B dense foundation model trained from scratch with extreme…
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement
arXiv:2609.13406v1 Announce Type: new Abstract: When we speak of recursive self-improvement (RSI), are we speaking of a phenomenon, a mechanism, or a…
Vibe Patenting: Evaluating LLM Judges for Professional Patent-Drafting Agents
arXiv:2609.13422v1 Announce Type: new Abstract: LLM judges are increasingly used to evaluate and improve AI-generated outputs, yet their reliability for…
CLSP-REQA: A Real-Time Quality-Aware Closed-Loop Seizure Prediction Framework with Mamba-BiLSTM and Confidence-Gated Intervention
arXiv:2606.00074v2 Announce Type: replace-cross Abstract: Reliable seizure prediction is a prerequisite for closed-loop neurostimulation therapy, yet…
AI Economist Agent: An Agentic Framework for Evidence-Based Economic and Financial Analysis with RAG, Knowledge Graphs, and Large Language Models
arXiv:2606.20041v2 Announce Type: replace-cross Abstract: We propose an AI economist agent for economic and financial scenario analysis. Scenario design…
From Agent Traces to Trust: A Survey of Evidence Tracing and Execution Provenance in LLM Agents
arXiv:2606.04990v5 Announce Type: replace-cross Abstract: Large language model (LLM)-based agents are evolving from passive text generators into…
LLM-Ideoplasticity: Measuring Ideological Plasticity in the Political Behavior of LLMs as a Context-Conditioned Distribution
arXiv:2606.28335v3 Announce Type: replace-cross Abstract: We argue, with systematic empirical evidence, that a large language model’s political ideology…
Adaptive Perturbation Selection for Contrastive Audio Decoding
arXiv:2607.00247v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) frequently hallucinate by overriding acoustic evidence with…
Causal Past Logic for Runtime Verification of Distributed LLM Agent Workflows
arXiv:2605.20923v2 Announce Type: replace-cross Abstract: We study runtime monitoring for distributed LLM-agent workflows. In an asynchronous execution, a…
