arXiv:2609.18461v2 Announce Type: replace Abstract: Personalized agents are required to reason over long-term history interactions to infer both explicit…
Category: AI
Prices go up in 7 days — get your Disrupt ticket now
Current ticket pricing ends September 25 at 11:59 p.m. PT. Join 10,000+ founders, investors, and tech leaders at Disrupt and save up to $200 on your…
A Unified Evaluation Framework for Trustworthy Large Language Models, Agentic AI, and Multimodal Systems
arXiv:2609.19524v2 Announce Type: replace Abstract: Benchmark scores alone provide an incomplete basis for assessing the trustworthiness of modern…
Planetary Prediction Engine: Autonomous Geospatial Prediction via Intelligent Data Selection and Foundation Model Embeddings
arXiv:2608.26088v2 Announce Type: replace Abstract: Addressing critical global challenges, from food security and disaster risk to disease outbreaks and…
A visual large language foundational model for medical image recognition using clinician-contributed online resources
arXiv:2609.06914v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated strong capabilities across diverse domains, showing…
Fraglingo: Molecular Design via Attachment-Aware Autoregressive Fragment Generation
arXiv:2609.13519v2 Announce Type: replace Abstract: We introduce Fraglingo, an autoregressive molecular generator that constructs molecules step by step…
Building standards for the next phase of AI
OpenAI outlines a path to shared global AI standards, calling for coordinated evaluation, reporting, and governance to improve safety.
Balance of Benchmarks: Semantic Density Reweighting for Task-Conditioned Model Comparison
arXiv:2608.30044v3 Announce Type: replace Abstract: Model comparison increasingly relies on large collections of publicly reported benchmark scores, yet…
Expanding OpenAI Academy with new learning paths
Explore new OpenAI Academy learning paths for employees, developers, leaders, educators, and students to build and demonstrate practical AI skills.
MOSCOPT: Mixture-of-Skills Collective Optimization for LLM Agents
arXiv:2609.14399v2 Announce Type: replace Abstract: Natural language prompts and skills serve as the strategic backbone of LLM-based agents. Recent…
