arXiv:2604.27167v3 Announce Type: replace-cross Abstract: On the named Prisoner’s Dilemma under direct prompting, three larger instruction-tuned models,…
Category: AI
Why Do LLMs Struggle in Strategic Play? Broken Links Between Observations, Beliefs, and Actions
arXiv:2605.00226v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly tasked with strategic decision-making under…
Rhamba: Region-Aware Hybrid Attention-Mamba Framework for Self-Supervised Learning in Resting-State fMRI
arXiv:2605.01240v3 Announce Type: replace-cross Abstract: Self-supervised pretraining is promising for large-scale neuroimaging, yet the impact of…
Representation Before Training: A Practical Benchmark for Generative Medical Event Model Tokenization
arXiv:2604.16775v3 Announce Type: replace-cross Abstract: Generative medical event models use tokenized sequences of patient timelines as input, but…
Evolving Skill Modules under a Fixed Planner: Versioning, Rollback, and Runtime Governance for Long-Lived Robot Systems
arXiv:2604.07799v3 Announce Type: replace-cross Abstract: Robots deployed for long periods keep improving their skills, and each update changes a released…
How do LLMs Compute Verbal Confidence
arXiv:2603.17839v4 Announce Type: replace-cross Abstract: Verbal confidence — prompting LLMs to state their confidence as a number or category — is…
Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation
arXiv:2603.16086v2 Announce Type: replace-cross Abstract: While recent Vision-Language-Action (VLA) models have begun to incorporate audio, they typically…
OpenAI forms math advisory group as its AI resolves more than 100 open problems
The group won’t be given leeway to slow down or redirect OpenAI’s ongoing mathematical research.
Lessons Without Borders? Evaluating Cultural Alignment of LLMs Using Multilingual Story Moral Generation
arXiv:2604.08797v2 Announce Type: replace-cross Abstract: Stories are key to transmitting values across cultures, but their interpretation varies across…
Taming the Adversary: A Cost-to-Disturbance Ratio Approach to Adversarial Reinforcement Learning
arXiv:2603.12110v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) policies trained in simulation often degrade once deployed on real…
