arXiv:2608.16349v1 Announce Type: new Abstract: Large language model (LLM) agents may assist flight crews with complex decisions and task execution, but…
WhiteFiber Proposes $250M Convertible Senior Notes to Fund Data Center Expansion
WhiteFiber, the AI infrastructure and high-performance computing provider, said on August 18, 2026 that it intends to raise $250 million through a private…
DriveCache: Action-Aware Caching for Driving World Model Inference
arXiv:2608.16354v1 Announce Type: new Abstract: Driving video generation models support autonomous-driving development by predicting controllable future…
BaT: Towards Self-Evolving Medical Research Agent with Stage Rubrics
arXiv:2608.16211v1 Announce Type: new Abstract: Long-horizon agents are beginning to automate complete workflows that produce code, reports, and research…
Beyond Asking: A Pipeline for Personalized Game Generation that Reads Players from Behavior
arXiv:2608.16196v1 Announce Type: new Abstract: Personalized game generation requires inferring a player’s abilities and behavioral style from how they…
Trajectory-Level Automatic Curriculum Learning for Legged Locomotion on Unstructured Terrain
arXiv:2608.16164v1 Announce Type: new Abstract: Training locomotion policies for complex unstructured terrain requires a curriculum to avoid early…
Baseline-Relative Counterfactual Refinement for Bit-Aware Visual Token Communication
arXiv:2608.16192v1 Announce Type: new Abstract: Generative visual-token communication reduces transmission load by sending only selected discrete tokens…
Strengthening Democratic Oversight in National Security
OpenAI launches an initiative to strengthen democratic oversight of AI in national security, supporting government institutions with tools, training, and…
Competing at Every Price Point with Agentic Evolution over a Menu of LLMs
arXiv:2608.16207v1 Announce Type: new Abstract: Consider a firm that surveys its competition for a particular agentic task and seeks to offer superior…
AI News Brief Hourly Summary 2026-08-18 22h : 15 posts
15 posts published in the last hour 19:32Assessing LLMs’ mathematical abilities requires understanding the various mechanisms of mathematical creativity 19:32FeatureHospital: A Skill-Driven Multi-Agent Framework for Automated Algorithm Customization in Multi-View Multi-Label Feature Selection 19:32When Single-Dataset Conclusions Fail: A 45-Task Study…
Assessing LLMs’ mathematical abilities requires understanding the various mechanisms of mathematical creativity
arXiv:2608.16118v1 Announce Type: new Abstract: How should we assess whether large language models can perform mathematical invention? I argue that this…
FeatureHospital: A Skill-Driven Multi-Agent Framework for Automated Algorithm Customization in Multi-View Multi-Label Feature Selection
arXiv:2608.16148v1 Announce Type: new Abstract: Multi-view multi-label feature selection aims to identify a compact and informative feature subset from…
When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification
arXiv:2608.16147v1 Announce Type: new Abstract: Class-imbalance handling is routinely evaluated on a single benchmark dataset, and the resulting…
Protein Structure Prediction: From Evolutionary Constraints to Generative Modeling
arXiv:2608.16094v1 Announce Type: new Abstract: Accurate protein structure prediction is fundamental to structural biology because protein structure…
Gemini in Chrome Opens to All U.S. Android Users as Auto Browse Goes Mobile
Google opened Gemini in Chrome to all Android users in the United States on August 18, 2026, bringing its built-in browsing assistant to the full U.S.…
TRCA: Transition-wise Rubric Credit Assignment for Long-horizon LLM Agents
arXiv:2608.16156v1 Announce Type: new Abstract: Long-horizon large language model (LLM) agents are typically optimized with sparse terminal outcomes,…
ALPS: Measuring Valid Creativity in Large Language Models with Mathematical Construction
arXiv:2608.15979v1 Announce Type: new Abstract: Large language models produce outputs presented as discoveries – new proofs, conjectures, or molecules.…
Governance at the Boundary: How Agent Decomposition Degrades Policy Compliance
arXiv:2608.16055v1 Announce Type: new Abstract: Existing agent benchmarks ask whether the agent finished the task. We ask whether it finished it within…
