arXiv:2606.06114v3 Announce Type: replace Abstract: Self-evolving agents improve through continual self-play and self-generated learning signals, but…
Tag: cs.AI updates on arXiv.org
Measuring Curriculum Alignment across Topical Coverage, Competency, and Cognitive Depth: A Longitudinal Framework Applied to CS2013 and CS2023
arXiv:2606.19469v2 Announce Type: replace Abstract: Undergraduate computer science is governed by international curricular guidelines revised about once a…
Shielded Analysis: Certification and Characterization of Defensibility in Systems under Adversarial Interaction
arXiv:2606.13621v2 Announce Type: replace Abstract: Formal safety analysis determines whether a system admits a safe defense; adaptive evaluation…
Superficial Beliefs in LLM Decision-Making
arXiv:2606.11016v2 Announce Type: replace Abstract: We ask whether large language models (LLMs) merely imitate rationales when choosing between two…
Rhythm of the Deep: Two-tier acoustic organization of sperm-whale codas from click waveforms to second-order sequence dependence
arXiv:2606.16084v3 Announce Type: replace Abstract: Sperm-whale codas are conventionally characterized by click count and inter-click intervals (ICIs),…
How Clinicians Think and What AI Can Learn From It
arXiv:2601.12547v2 Announce Type: replace Abstract: Clinical artificial intelligence increasingly builds high-dimensional representations of patients, yet…
Same Answer, Different Representations: Hidden instability in VLMs
arXiv:2602.06652v2 Announce Type: replace Abstract: The robustness of Vision Language Models (VLMs) is commonly assessed through output-level invariance,…
FactorEngine: A Program-level Knowledge-Infused Factor Mining Framework for Quantitative Investment
arXiv:2603.16365v3 Announce Type: replace Abstract: We study alpha factor mining, the automated discovery of predictive signals from noisy, non-stationary…
Autonomous Assessment of Generalizability of AI Agent Capabilities
arXiv:2512.16733v4 Announce Type: replace Abstract: Safe deployment of black-box AI (BBAI) systems such as foundation model agents requires methods for…
Multi-Agent Collaboration for Automated Design Exploration on High Performance Computing Systems
arXiv:2603.11515v2 Announce Type: replace Abstract: Today’s scientific challenges, from climate modeling to Inertial Confinement Fusion design to novel…
