arXiv:2608.24979v2 Announce Type: replace Abstract: Scientific agents increasingly analyze data, execute code, and produce research artifacts, yet most…
One week left to book your exhibit table at TechCrunch Disrupt 2026
Only one week left to secure your exhibit table. Tables are limited and can sell out before the September 18 deadline.
A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation
arXiv:2609.01315v2 Announce Type: replace Abstract: Building an omni-modal foundation model means evaluating it across text, image, video, and audio.…
OpenAI’s feud with mathematicians is only escalating
Twenty-five leading mathematicians signed an open letter arguing that AI labs are threatening their intellectual work.
From Monolithic Blending to Agentic Orchestration: Dynamic Response for Conversational Assistants at Scale
arXiv:2609.05758v2 Announce Type: replace Abstract: Conversational assistants can blend retrieval, action selection, escalation, and wording in a single…
Y Combinator’s Garry Tan wants U.S. open-weight AI labs to ‘distill’ frontier models, too
Tan argues that frontier models themselves trained on public human knowledge so access to capable AI should be “a form of public good.”
Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation
arXiv:2609.04298v3 Announce Type: replace Abstract: Evaluating agents on the growing number of agentic benchmarks is challenging because they often…
AI News Brief Hourly Summary 2026-09-11 23h : 16 posts
16 posts published in the last hour 20:32LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph Forecasting 20:32Physics of Agents: Statistical Mechanics Predicts Collective Behavior of AI Agents 20:32Roundtables: Will AI really kill us all? 20:32A Human Audit of OpenAIs…
LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph Forecasting
arXiv:2608.06765v2 Announce Type: replace Abstract: Continuous-time dynamic graph models predict future links by compressing past interactions into neural…
Physics of Agents: Statistical Mechanics Predicts Collective Behavior of AI Agents
arXiv:2608.16578v2 Announce Type: replace Abstract: AI agents increasingly operate as part of interacting systems rather than in isolation. As agents…
Roundtables: Will AI really kill us all?
Employees at the world’s leading AI labs are saying there’s a real possibility that advanced AI could destroy humanity. Are they right? Or is this more…
A Human Audit of OpenAIs AI-Generated Mathematical Proofs
arXiv:2608.14673v3 Announce Type: replace Abstract: We assess 18 chapter-specific reviews of the ten mathematical results announced by OpenAI on 1 August…
Final, final, final call for TechCrunch Disrupt 2026 Side Events
The absolute last chance to apply to host an official Side Event during TechCrunch Disrupt 2026 is tonight, September 11, at 11:59 p.m. PT.
Learning to Predict Middle-Layer Attention in MLLMs for Visual Token Pruning
arXiv:2608.06411v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) achieve strong performance across diverse vision-language…
Kimi-maker Moonshot AI targets $2B in annual revenue
While K3’s usage figures have declined slightly in recent months, OpenRouter data currently shows as many as 300 billion tokens being generated each day…
Dear Algo: A Precision-First Agentic Intent Layer for Unified Search and Recommendation
arXiv:2608.15877v3 Announce Type: replace Abstract: Search and recommendation serve a shared discovery objective but encode intent differently. We study…
ViSR-KGC: Visual Subgraph Reasoning with Vision-Language Models for Multimodal Knowledge Graph Completion
arXiv:2608.05833v3 Announce Type: replace Abstract: Knowledge graph completion (KGC) aims to infer missing entities or relations from incomplete graph…
Self-Evolving Scientific Agent Designs Physically Reasoned White-Box Fluid Control
arXiv:2606.08405v4 Announce Type: replace Abstract: While neural networks excel in autonomous control, their black-box nature makes control decisions…
