The deal is less a new logo than a deepening of an existing relationship. The scope spans the properties fans actually touch: FIFA ID as a unified…
When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning
arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We…
Blacksmith Raises Funding for AI Code Validation Layer
Blacksmith, a San Francisco startup that runs continuous integration workloads on purpose-built hardware, has raised a $45 million Series B led by Peak XV…
“YES! YES! I absolutely love this insight!” Affirmative Narration as Interactional Strategy in Dialogues with LLM Chatbots
arXiv:2607.28646v2 Announce Type: cross Abstract: This article analyses narrative mechanisms that are common in dialogues with LLM chatbots. In…
AI News Brief Hourly Summary 2026-08-12 14h : 15 posts
15 posts were published in the last hour 11:33 : Long-Horizon AI Research for Grothendieck Constant: A Case Study in Human-AI Mathematical Collaboration 11:33 : Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding 11:33 : Virgin Atlantic sharpens…
Long-Horizon AI Research for Grothendieck Constant: A Case Study in Human-AI Mathematical Collaboration
arXiv:2608.11195v1 Announce Type: new Abstract: AI agents are increasingly used in mathematics research, but it is often unclear how to use them…
Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding
arXiv:2608.11095v1 Announce Type: new Abstract: Agentic coding READMEs like CLAUDE.md grow without bound in real repositories, stopping only when the…
Virgin Atlantic sharpens customer journeys with ChatGPT Work
Virgin Atlantic is accelerating research, product planning, and decision-making with ChatGPT Work, helping teams connect signals across the customer…
The Gaussian-Multinoulli Restricted Boltzmann Machine: A Potts Model Extension of the GRBM
arXiv:2505.11635v2 Announce Type: cross Abstract: Many real-world tasks, from associative memory to symbolic reasoning, benefit from discrete, structured…
AI code-testing startup Blacksmith’s valuation jumps almost 10x in less than a year
Blacksmith says revenue has grown more than tenfold over the past year.
RTSKG: Building a Rail Transit Station Knowledge Graph Dataset
arXiv:2608.11080v1 Announce Type: new Abstract: Rail transit systems play a vital role in urban mobility and economic development. As key components of…
How Zapier transformed core marketing processes with ChatGPT Work
The enterprise marketing team at Zapier uses ChatGPT Work to reduce the number of drop-offs in its lead funnel, build campaign assets, and automate…
sLTN: Structural Logic Tensor Networks
arXiv:2608.11136v1 Announce Type: new Abstract: Logic Tensor Networks (LTN) provide a neurosymbolic framework in which first-order logic is interpreted…
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
arXiv:2608.11079v1 Announce Type: new Abstract: Self-evolving agents accumulate reusable skills by appending successful procedures and failure fixes. Over…
XCoT-VLA: Executable Chain-of-Thought for Vision-Language-Action Driving
arXiv:2608.10976v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can connect scene understanding, semantic reasoning, and trajectory…
V-FiLLM: Verified Financial LLM Reasoning Benchmark
arXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial…
ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling
arXiv:2608.10928v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by allocating additional inference-time compute to…
Pakistani Judges Give Their Verdict on JudgeGPT
Judges around the world have made headlines for illicitly using generative AI in their work. But in Pakistan, a large-scale trial of a specially designed…
