ChatGPT Ads has hit $1 billion in annualised revenue run rate in under 200 days, and OpenAI is expanding self-service ads to new regions. Tens of…
trajectory-judge: What Outcome-Only LLM Judges Miss on Agent Trajectories
arXiv:2609.00038v1 Announce Type: cross Abstract: Outcome-only evaluation is the production default for LLM agents: show a judge the request and the final…
EvoSCM: Scientific Belief Revision Through Causal Model Evolution and Experimentation
arXiv:2609.01526v1 Announce Type: new Abstract: Scientific agents must learn not only how to reason, but also what to believe. However, existing LLM…
Proactive cyber defense for governments and enterprises
Introducing Fairwind Program
When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation
arXiv:2609.01519v1 Announce Type: new Abstract: Interactive simulations increasingly evaluate policies in markets populated by language-model agents.…
India’s richest man now wants to turn aging computers into AI-ready PCs
Jio is betting it can turn an aging computer into an AI-ready PC for as little as about $11 for two months.
Can LLMs Discover Scientific Laws in Real and Parallel Worlds?
arXiv:2609.01552v1 Announce Type: new Abstract: Scientific equation discovery has long been central to scientific progress, proceeding through iterative…
Huskeys Raises $27M Series A to Build the Security Control Layer for the AI Driven Network
Huskeys has raised $27 million in Series A funding, led by Blackstone Innovations Investments, as the cybersecurity startup looks to address a growing…
InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information
arXiv:2608.29632v1 Announce Type: cross Abstract: Competitive programming is increasingly being used to evaluate the algorithmic reasoning capabilities of…
Quantifying User Behavior Patterns to Build Better Predictive Features
Simply knowing that a 35-year-old male in Seattle clicked 12 times last month tells you almost nothing about his intent.
Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers
arXiv:2609.01567v1 Announce Type: new Abstract: Vision-Language Models (VLMs) provide useful priors for interactive decision-making, but using them…
AI News Brief Hourly Summary 2026-09-02 18h : 24 posts
24 posts published in the last hour 15:33Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations 15:33EdiTikZ: Scientific Figure Editing from Revision Trajectories 15:33Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers 15:33HiddenLayer Raises $100M…
Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations
arXiv:2609.01408v1 Announce Type: new Abstract: A fundamental challenge in artificial intelligence is the transformation of observations into explicit…
EdiTikZ: Scientific Figure Editing from Revision Trajectories
arXiv:2609.01409v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance in generating scientific figures from text or…
Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers
arXiv:2609.01466v1 Announce Type: new Abstract: A long-horizon agent’s trace outgrows both of its consumers: the human observer monitoring the run, and…
HiddenLayer Raises $100M Series B to Expand AI Agent Security Platform
HiddenLayer, an Austin-based AI security company, announced on September 2, 2026 that it has raised a $100 million Series B round led by Delta-v Capital,…
Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement
arXiv:2609.01481v1 Announce Type: new Abstract: This paper studies autonomous software development, in which LLM-based coding agents transform high-level…
OpenAI supports California’s bill to advance youth AI safety
OpenAI supports California SB 1119, advancing strong, age-appropriate AI safeguards for teens while preserving opportunities to learn, create, and explore.
