OpenAI previewed the precautions it is taking as it prepares to release Astra, its newest, cyber-critical LLM.
Tag: AI
Cross-Regional Grapevine Cold Hardiness Prediction via Learned Multimodal Latent Representations
arXiv:2608.31097v1 Announce Type: new Abstract: Accurate daily predictions of cold hardiness in woody plants are critical in regions where freezing…
Measure Before You Manage: Evaluating Agent Working Memory in Coding Agents
arXiv:2608.31057v1 Announce Type: new Abstract: Agent working memory is heterogeneous. Objects such as instructions, artifacts, tool outputs, and…
Learning Action Models with Conditional and Quantified Effects via Uncertainty-Guided Exploration
arXiv:2608.30955v1 Announce Type: new Abstract: Accurate action models are critical for effective planning. Existing action-model learning methods largely…
Anthropic opens Claude AI text detection to regulators, media, fact-checkers, and others
Anthropic is launching an API that lets regulators, media outlets, and researchers check whether text carries Claude’s digital watermark. The EU AI Act…
MNIST-PRO: MNIST is Back as a Partially Observable World for AI Agents
arXiv:2608.31022v1 Announce Type: new Abstract: AI agents in partially observable environments need to coordinate active sensing with working memory to…
The latest AI news we announced in August 2026
Transitioning cards: 1. Text “Gemini 3.7 Flash” next to the Gemini logo icon; 2. a photo of a pixel phone; 3. Google Gemini logo above the text “Claim…
Wrong Prediction, Right Answer: Recovering Evidence from Collapsed LLM Sequence Scores
arXiv:2608.31068v1 Announce Type: new Abstract: When a large language model fails a reasoning task, it is often assumed to lack the underlying capability.…
Google’s Android update tackles motion sickness, accessibility, and more
The new features coming to Android are aimed at reducing motion sickness, helping blind users navigate their surroundings, and more.
Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence
arXiv:2608.31075v1 Announce Type: new Abstract: Recent advances in large reasoning models (LRMs) have shown that reinforcement learning with verifiable…
