arXiv:2609.09126v1 Announce Type: new Abstract: Amari’s contributions to information geometry and machine learning are well known. Here, we revisit…
Author: script
ExecCritic: Learn to Test, Test to Improve for Coding Agents
arXiv:2609.09133v1 Announce Type: new Abstract: Execution feedback can guide coding agents toward correct repository repairs, but only when the tests…
Deposon: An Auditable, Conservation-Guaranteed, Game-Theoretically Tested Scattering Layer over LLM Reasoning Paths
arXiv:2609.09001v1 Announce Type: new Abstract: Multi-step LLM reasoning lacks a machine-recheckable ledger: discarded reasoning paths leave no auditable…
Time-Varying Data as Sheaves: an Invitation to Narratives
arXiv:2609.09056v1 Announce Type: new Abstract: Modern science and engineering increasingly rely on time-varying data, yet the mathematical tools used to…
The Surprising Effectiveness of Approximate Value Iteration in Self-Play
arXiv:2609.09094v1 Announce Type: new Abstract: Combining search with function approximation has driven major advances in game-playing programs, making…
Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning
arXiv:2609.09030v1 Announce Type: new Abstract: Chain-of-thought reasoning provides a structured computation between a model’s input and final answer. Yet…
Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster
Production LLM applications rarely receive a question nobody has asked before. Support assistants and RAG pipelines field the same intents thousands of…
Everything in Moderation: Per-Domain Coverage Optima and Alignment-Resistant Domain Gaps in Multi-Domain Mid-Training
arXiv:2609.09081v1 Announce Type: new Abstract: Mid-training, the stage between pre-training and alignment, is where a model’s per-domain data composition…
AI News Brief Hourly Summary 2026-09-11 01h : 16 posts
16 posts published in the last hour 22:33PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving 22:33SkillAdam: Stable and Efficient Skill Evolution for Agents 22:33API Benchmark Scores Do Not Reliably Transfer to Chatbot Interfaces 22:33Good Pretraining, Bad…
PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving
arXiv:2609.08965v1 Announce Type: new Abstract: Ensuring the safety of autonomous driving is a critical challenge. Scenario-based testing is a systematic…
