arXiv:2109.08813v1 Announce Type: cross Abstract: Seismic wave velocity of underground rock plays important role in detecting internal structure of the…
Neural Symbollic Regression Using Deep Learning and Sparse Modelling
arXiv:2609.01102v1 Announce Type: cross Abstract: Symbolic Regression (SR) seeks to find succinct mathematical expressions that represent the fundamental…
Convergence issues in Relational Concept Analysis based on AOC-posets
arXiv:2609.00054v1 Announce Type: cross Abstract: Formal Concept Analysis (FCA) is an approach for conceptual classification building and rule discovery…
AI News Brief Hourly Summary 2026-09-11 02h : 12 posts
12 posts published in the last hour 23:32MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents 23:32SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research? 23:32Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where…
MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents
arXiv:2609.09115v1 Announce Type: new Abstract: Long horizon Large Language Model (LLM) agents rely on external memory systems to preserve user…
SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
arXiv:2609.09113v1 Announce Type: new Abstract: While research on recursive self-improvement (RSI) has predominantly automated model training pipelines,…
Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails
arXiv:2609.09134v1 Announce Type: new Abstract: Agent harnesses (the system prompt, tool set, execution hooks, and context-management scaffolding around a…
A Generalization of Amari’s Bayesian Duality
arXiv:2609.09126v1 Announce Type: new Abstract: Amari’s contributions to information geometry and machine learning are well known. Here, we revisit…
ExecCritic: Learn to Test, Test to Improve for Coding Agents
arXiv:2609.09133v1 Announce Type: new Abstract: Execution feedback can guide coding agents toward correct repository repairs, but only when the tests…
Deposon: An Auditable, Conservation-Guaranteed, Game-Theoretically Tested Scattering Layer over LLM Reasoning Paths
arXiv:2609.09001v1 Announce Type: new Abstract: Multi-step LLM reasoning lacks a machine-recheckable ledger: discarded reasoning paths leave no auditable…
Time-Varying Data as Sheaves: an Invitation to Narratives
arXiv:2609.09056v1 Announce Type: new Abstract: Modern science and engineering increasingly rely on time-varying data, yet the mathematical tools used to…
The Surprising Effectiveness of Approximate Value Iteration in Self-Play
arXiv:2609.09094v1 Announce Type: new Abstract: Combining search with function approximation has driven major advances in game-playing programs, making…
Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning
arXiv:2609.09030v1 Announce Type: new Abstract: Chain-of-thought reasoning provides a structured computation between a model’s input and final answer. Yet…
Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster
Production LLM applications rarely receive a question nobody has asked before. Support assistants and RAG pipelines field the same intents thousands of…
Everything in Moderation: Per-Domain Coverage Optima and Alignment-Resistant Domain Gaps in Multi-Domain Mid-Training
arXiv:2609.09081v1 Announce Type: new Abstract: Mid-training, the stage between pre-training and alignment, is where a model’s per-domain data composition…
AI News Brief Hourly Summary 2026-09-11 01h : 16 posts
16 posts published in the last hour 22:33PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving 22:33SkillAdam: Stable and Efficient Skill Evolution for Agents 22:33API Benchmark Scores Do Not Reliably Transfer to Chatbot Interfaces 22:33Good Pretraining, Bad…
PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving
arXiv:2609.08965v1 Announce Type: new Abstract: Ensuring the safety of autonomous driving is a critical challenge. Scenario-based testing is a systematic…
SkillAdam: Stable and Efficient Skill Evolution for Agents
arXiv:2609.08944v1 Announce Type: new Abstract: Agent skills provide a lightweight way to equip frozen language-model agents with domain knowledge and…
