arXiv:2603.11515v2 Announce Type: replace Abstract: Today’s scientific challenges, from climate modeling to Inertial Confinement Fusion design to novel…
Tag: AI
MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning
arXiv:2505.24846v3 Announce Type: replace Abstract: Reward modeling is a key step in building safe foundation models when applying reinforcement learning…
TrafficGamer: Reliable and Flexible Traffic Simulation for Safety-Critical Scenarios with Game-Theoretic Oracles
arXiv:2408.15538v4 Announce Type: replace Abstract: While modern Autonomous Vehicle (AV) systems can develop reliable driving policies under regular…
EVINCE: Optimizing Multi-LLM Dialogues Using Conditional Statistics and Information Theory
arXiv:2408.14575v5 Announce Type: replace Abstract: EVINCE (Entropy and Variation IN Conditional Exchanges) is a novel framework for optimizing multi-LLM…
Fairness at Every Intersection: Uncovering and Mitigating Intersectional Biases in Multimodal Clinical Predictions
arXiv:2412.00606v2 Announce Type: replace Abstract: Biases in automated clinical decision-making using Electronic Healthcare Records (EHR) impose…
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
arXiv:2505.23686v4 Announce Type: replace Abstract: Learning to collaborate with previously unseen partners is a fundamental generalization challenge,…
LACE: Layer-Wise Compression for Dynamic Frame Rate Codecs
arXiv:2609.17509v1 Announce Type: cross Abstract: Neural audio codecs are a key component in speech language modeling. However, their high frame rates…
Agentic Societies Need a Social Harness
arXiv:2609.17527v1 Announce Type: cross Abstract: An agentic society is a collection of AI agents that coordinate autonomously across trust boundaries, on…
ENCP: Episode-Normalized Conformal Prediction for Vision-and-Language Navigation
arXiv:2609.17499v1 Announce Type: cross Abstract: Uncertainty estimation for Vision-Language-Navigation (VLN) models is a critical task since it can help…
When Should LLMs Abstain? Chain-of-Self-Questioning for Selective Risk Control
arXiv:2609.17516v1 Announce Type: cross Abstract: Large language models can produce fluent answers when their factual support is weak. This paper…
