arXiv:2602.18532v3 Announce Type: replace-cross Abstract: Following the rise of large foundation models, Vision-Language-Action models (VLAs) emerged,…
Category: AI
ICA: Information-Aware Credit Assignment for Visually Grounded Long-Horizon Information-Seeking Agents
arXiv:2602.10863v2 Announce Type: replace-cross Abstract: Long-horizon reinforcement learning for information seeking agents remains difficult because…
PatientHub: A Unified Framework for Patient Simulation
arXiv:2602.11684v2 Announce Type: replace-cross Abstract: As Large Language Models increasingly power role-playing applications, simulating patients has…
Anytime Pretraining: Horizon-Free Learning-Rate Schedules with Weight Averaging
arXiv:2602.03702v2 Announce Type: replace-cross Abstract: Large language models are increasingly trained in continual or open-ended settings, where the…
Liquid AI Open-Sources Pipette: A Reproducible Benchmarking Suite That Measures On-Device Models, Quantization, Runtime and Hardware Together
Model cards report quality under server-class, full-precision conditions. Those numbers rarely predict how the same model behaves on a phone. This week,…
You Can Learn Tokenization End-to-End with Reinforcement Learning
arXiv:2602.13940v3 Announce Type: replace-cross Abstract: Tokenization is a hardcoded compression step which remains in the training pipeline of Large…
Seeing vs. Believing: Evaluating the Language Bias of Open-Source MLLMs in Counter-Intuitive Scenes
arXiv:2601.07737v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable performance in mainstream…
Minimal Decision Dynamics and Contextual Probability: A Quantum Tug-of-War Model
arXiv:2601.10034v3 Announce Type: replace-cross Abstract: Decision making often exhibits context dependence that is difficult to accommodate within a…
Ad Insertion in LLM-Generated Responses
arXiv:2601.19435v2 Announce Type: replace-cross Abstract: Sustainable monetization of large language models (LLMs) remains a critical open challenge.…
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
arXiv:2601.16520v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in visual recognition…
