arXiv:2609.13731v1 Announce Type: new Abstract: The transition from passive foundation models to autonomous, goal-directed agentic AI systems has…
Tag: cs.AI updates on arXiv.org
JaxAHT: A JAX-Based Library for Ad Hoc Teamwork
arXiv:2609.13716v1 Announce Type: new Abstract: Ad Hoc Teamwork (AHT) addresses the challenge of designing agents capable of coordinating with novel…
MANAS-2: Constrained Reconstruction for EEG Foundation Models
arXiv:2609.13717v1 Announce Type: new Abstract: Masked reconstruction is widely used for EEG foundation models, but optimizing reconstruction on low-SNR…
How Many Thoughts Can a Vector Hold? The Capacity of Reasoning by Superposition
arXiv:2609.13747v1 Announce Type: new Abstract: Large language models solve hard problems through intermediate computations across multi-step reasoning.…
Drift-Constrained Optimization: Only Direction Matters in Fine-Tuning Instruct Models
arXiv:2609.13680v1 Announce Type: new Abstract: Fine-tuning instruct models often improves target performance while inducing behavioral drift from the…
Degraded but Not Entirely Ineffective: PE-Based Deformable Graph Neural Networks
arXiv:2609.13712v1 Announce Type: new Abstract: Many real-world scenarios can be represented using graph-structured data. However, traditional GNNs that…
Windowed A-K-MDP
arXiv:2609.13676v1 Announce Type: new Abstract: Markov decision processes (MDPs) are used to support decision-making in conservation of biodiversity, but…
Recoverability as a System Primitive for Long-Horizon AI Agents
arXiv:2609.13672v1 Announce Type: new Abstract: AI agents can be interrupted while editing files, calling tools, or carrying out multi-step tasks.…
Enhancing Event Candidate Acquisition for Event Linking
arXiv:2609.13670v1 Announce Type: new Abstract: Event linking associates event mentions in text with entries in a knowledge base (KB), or identifies them…
Safety as a Constraint: Fine-Tuning a LLM Recommender to Explain Itself
arXiv:2609.13657v1 Announce Type: new Abstract: Traditional recommender systems are typically trained to predict what item users will interact with next,…
