arXiv:2602.06652v2 Announce Type: replace Abstract: The robustness of Vision Language Models (VLMs) is commonly assessed through output-level invariance,…
Category: AI
FactorEngine: A Program-level Knowledge-Infused Factor Mining Framework for Quantitative Investment
arXiv:2603.16365v3 Announce Type: replace Abstract: We study alpha factor mining, the automated discovery of predictive signals from noisy, non-stationary…
House Passes Ratepayer Protection Act on Data Center Power Costs
The U.S. House passed the Ratepayer Protection Act on September 16, 2026, voting 417 to 3 to require state utility regulators to consider standards that…
Autonomous Assessment of Generalizability of AI Agent Capabilities
arXiv:2512.16733v4 Announce Type: replace Abstract: Safe deployment of black-box AI (BBAI) systems such as foundation model agents requires methods for…
OpenAI, Anthropic, Google have been in talks on AI safety for weeks
OpenAI confirms weeks of AI safety talks with Anthropic and Google DeepMind, as Trump’s team dismisses safety concerns and pushes to keep pace with China.
Multi-Agent Collaboration for Automated Design Exploration on High Performance Computing Systems
arXiv:2603.11515v2 Announce Type: replace Abstract: Today’s scientific challenges, from climate modeling to Inertial Confinement Fusion design to novel…
MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning
arXiv:2505.24846v3 Announce Type: replace Abstract: Reward modeling is a key step in building safe foundation models when applying reinforcement learning…
TrafficGamer: Reliable and Flexible Traffic Simulation for Safety-Critical Scenarios with Game-Theoretic Oracles
arXiv:2408.15538v4 Announce Type: replace Abstract: While modern Autonomous Vehicle (AV) systems can develop reliable driving policies under regular…
EVINCE: Optimizing Multi-LLM Dialogues Using Conditional Statistics and Information Theory
arXiv:2408.14575v5 Announce Type: replace Abstract: EVINCE (Entropy and Variation IN Conditional Exchanges) is a novel framework for optimizing multi-LLM…
Fairness at Every Intersection: Uncovering and Mitigating Intersectional Biases in Multimodal Clinical Predictions
arXiv:2412.00606v2 Announce Type: replace Abstract: Biases in automated clinical decision-making using Electronic Healthcare Records (EHR) impose…
