arXiv:2608.15147v1 Announce Type: new Abstract: Machine intelligence has conquered the symbolic world but stalled at the physical one. The stall is…
As AI beats doctors, regulators shouldn’t force a human into the loop, JAMA piece says
An opinion piece in the medical journal JAMA argues that autonomous AI will soon outperform any doctor-AI team at medical reasoning tasks. The authors…
SkillCommit: Evolving Agent Skills through Behaviorally Validated Scope Expansion
arXiv:2608.15165v1 Announce Type: new Abstract: Large language model (LLM) agents can continually improve without parameter updates by converting…
DOJ probes Andreessen Horowitz over partners sitting on competing AI boards
Andreessen Horowitz is the focus of an antitrust probe by the US Justice Department. The charge is that the firm’s partners sit on the boards of competing…
Demographic Injection in Medical Language Models under Diversity, Equity, and Inclusion Prompts
arXiv:2608.15254v1 Announce Type: new Abstract: Clinical-AI guidance increasingly recommends prompting language models to reason with attention to…
AI News Brief Hourly Summary 2026-08-18 14h : 19 posts
19 posts published in the last hour 11:33Platform Adaptation Under Governance Interventions: Actor Best-Response Modeling and an External Public-Case Benchmark 11:33Introducing ChatGPT for Teens: Built for learning, backed by protections 11:33OpenAI launches a ChatGPT version built for teens 11:33OpenAI Launches…
Platform Adaptation Under Governance Interventions: Actor Best-Response Modeling and an External Public-Case Benchmark
arXiv:2608.15131v1 Announce Type: new Abstract: Digital platforms govern by changing rules: rankings, monetization thresholds, moderation standards,…
Introducing ChatGPT for Teens: Built for learning, backed by protections
ChatGPT for Teens helps teens learn, think critically, and use AI with confidence, with stronger built-in protections, healthy-use features, and…
OpenAI launches a ChatGPT version built for teens
OpenAI is shipping a version of ChatGPT tailored to users aged 13 to 17. The article OpenAI launches a ChatGPT version built for teens appeared first on…
OpenAI Launches ChatGPT for Teens With Default Protections and Learning Tools
OpenAI on August 18, 2026 launched ChatGPT for Teens, a dedicated version of its chatbot that automatically applies to users aged 13 to 17 (and to anyone…
Translating finite-domain integer constraint models to CP/SMT/ILP/PB/SAT solvers with CPMpy
arXiv:2608.15143v1 Announce Type: new Abstract: Constraint solving is a declarative approach for solving combinatorial satisfaction and optimization…
Partnering with CodeAI to prepare the first AI generation
OpenAI and CodeAI are partnering to help students build AI literacy, think critically about AI, and develop the skills to use and shape it responsibly.
ReForge: Keeping ABR Algorithms Never Finished with Verified Large Language Model Edits
arXiv:2608.15138v1 Announce Type: new Abstract: Designing an ABR algorithm for one network scenario takes an engineer months, and large language models…
Don’t Let Cybercriminals Score: What the World Cup Taught Us About Dodging Scams
The World Cup recently wrapped, and for fans, the excitement of a month’s worth of matches is still fresh. For cybercriminals, it meant something else: an…
Anatomy of a Quantized Agent: VRAM Stability and Forecasting in Code-Synthesis Agentic Workloads
arXiv:2608.15117v1 Announce Type: new Abstract: Analytical models of peak VRAM consumption for LLM inference decompose memory into weight-storage,…
Proctoring in the Age of AI, Why Exam Design Still Comes First
Future-proofing online exams amid rising demand and tightening privacy laws. The global online exam software market, valued at $9.4b in 2025, is expected…
ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language Models
arXiv:2608.15145v1 Announce Type: new Abstract: Large Language Models (LLMs) have been increasingly adopted in Text-to-SQL systems, yet SQL errors remain…
StateM: Reaching 95.3% Raw Accuracy, or a \$15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling
arXiv:2608.15089v1 Announce Type: new Abstract: Long-horizon agents can fail even when their underlying models can solve the constituent steps. They may…
