Anthropic published its second company-wide Risk Report on August 14, 2026, and the headline change is a one-word upgrade in the wrong direction: the…
GeoCache: Training-Free Acceleration of Multi-View Texture Diffusion via Geometric Delta Transport
arXiv:2608.13255v1 Announce Type: cross Abstract: Geometry-conditioned multi-view diffusion enables high-quality 3D texture generation, but its repeated…
OpenAI Tells Investors Enterprise Revenue Has Overtaken Its ChatGPT Consumer Business
OpenAI’s finance chief told shareholders on August 14, 2026 that the company’s enterprise operation now generates more revenue than its ChatGPT-led…
Self-Referential Induction Increases Response Instability Relative to Unresolvable and Verifiable Questions in Large Language Models
arXiv:2608.13258v1 Announce Type: cross Abstract: Self-referential prompting has been shown to reliably induce large language models to produce…
Anthropic Explains the Mechanics of Claude’s Text Watermark
Anthropic published a detailed account on August 14, 2026 of how the text watermark in future Claude models works, identifying it as a version of the…
CoverPrune: Coverage-Driven Token Pruning for 3D VLMs via Optimal Transport
arXiv:2608.13226v1 Announce Type: cross Abstract: While 3D Vision-Language Models (3D VLMs) have demonstrated remarkable spatial reasoning capabilities,…
AI News Brief Hourly Summary 2026-08-14 23h : 11 posts
11 posts published in the last hour 20:32Better Decomposition, Free Aggregation: A Synthesizer-Folding Framework for Multilingual Multi-Hop Question Answering 20:32TRAPSBench: Vision-Language Models Encode but Fail to Express Epistemic Restraint 20:32GEM: A Generative Embedding Model Bridging Reasoning and Retrieval 20:31NARU: A…
Better Decomposition, Free Aggregation: A Synthesizer-Folding Framework for Multilingual Multi-Hop Question Answering
arXiv:2608.13160v1 Announce Type: cross Abstract: Multilingual retrieval-augmented generation (mRAG) equips large language models with access to globally…
TRAPSBench: Vision-Language Models Encode but Fail to Express Epistemic Restraint
arXiv:2608.13167v1 Announce Type: cross Abstract: When visual evidence is occluded or chaotic, models should abstain. In this paper, we show that…
GEM: A Generative Embedding Model Bridging Reasoning and Retrieval
arXiv:2608.13200v1 Announce Type: cross Abstract: Modern LLMs excel at reasoning and instruction following, enabling users to express complex and diverse…
NARU: A Benchmark for NARrative Evolution and Cultural Nuance Understanding in Japanese Extreme Long Video
arXiv:2608.13210v1 Announce Type: cross Abstract: Long-form video understanding encompasses tasks that go beyond retrieving isolated events, including…
LipCache: A Local Inference Proxy with Certified Caching for Edge Image Classification Service
arXiv:2608.13144v1 Announce Type: cross Abstract: As edge-side vision services continue to expand toward low-latency, high-throughput scenarios, reducing…
LOB-ID: Evaluating Synthetic Market Data by Inception Distances
arXiv:2608.13082v1 Announce Type: cross Abstract: Generative models of limit orderbook (LOB) data have advanced rapidly, but their evaluation often…
Sampling Luck Masquerades as Allocation Gain: Auditing Test-Time Budget Allocation for Neural Combinatorial Optimization
arXiv:2608.13087v1 Announce Type: cross Abstract: Neural combinatorial optimization (NCO) solvers report the best of many sampled solutions per instance,…
EgoMonth: A Month-Level Egocentric Video Benchmark for Long-Term Spatiotemporal Memory
arXiv:2608.13113v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to substantial progress in video…
TEMPO: Makespan-Aware Expert-Parallel Load Balancing Across Memory- and Compute-Bound Regimes
arXiv:2608.13057v1 Announce Type: cross Abstract: In expert-parallel (EP) MoE serving, every layer synchronizes at the slowest GPU. Dispatchers balance…
How kids feel about AI, in their own words
When we set out to talk to kids about artificial intelligence, we thought we knew what we’d hear. We expected some to tell us they were using it to cheat…
LigBench: A Unified and Human-Aligned Benchmark for LLM-based Research Idea Generation
arXiv:2608.13136v1 Announce Type: cross Abstract: With the rapid advancement of large language models (LLMs), research idea generation has attracted…
