arXiv:2608.03020v2 Announce Type: replace Abstract: Parameter-efficient post-training reduces the number of trainable parameters, but still requires…
Category: cs.AI updates on arXiv.org
Recursive Synthesis for Long-Horizon Terminal Tasks
arXiv:2608.05466v2 Announce Type: replace Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing…
SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse
arXiv:2608.05204v2 Announce Type: replace Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata,…
Cross-Layer Interaction under Weight-Space Ablation: A Closed-Form Attention Jacobian Bound and a Test on a Real Pretrained Model
arXiv:2608.03629v2 Announce Type: replace Abstract: A companion paper studies when activation patching and weight-space ablation agree, inside an…
CourseGraph: Finding overlaps and differences in Computer Science courses across universities
arXiv:2608.05910v2 Announce Type: replace Abstract: Student mobility programs such as Erasmus+ enable students to take courses at other universities,…
Property-driven Causal Abstractions for Markov Decision Processes
arXiv:2607.26787v3 Announce Type: replace Abstract: Markov Decision Processes (MDPs) are widely used as decision-making models, commonly specified over…
Homebot: A Personal AI Agent for Conversational Home Assistance and Automation
arXiv:2608.02254v2 Announce Type: replace Abstract: \texttt{Homebot} is a locally deployable AI agent for conversational household assistance and…
Can AI agents conduct open-ended AI research? Early evidence from two case studies
arXiv:2607.27191v2 Announce Type: replace Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether…
H+ Embedding: Harmonizing Global and Token-Level Retrieval with Context-Dependent Phrases
arXiv:2608.00065v3 Announce Type: replace Abstract: Terminology-intensive retrieval, especially in medical settings, depends on preserving multi-word…
OpenForgeRL: Train Harness-native Agents in Any Environment
arXiv:2607.21557v3 Announce Type: replace Abstract: Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to…