arXiv:2609.03667v1 Announce Type: cross Abstract: Generalising to unseen tasks remains a fundamental challenge in offline multi-agent reinforcement…
Author: script
Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners
arXiv:2609.03660v1 Announce Type: cross Abstract: The dominance of Neural Networks (NNs) in RL is partially due to their incremental learning capability,…
Doesn’t Stop Reasoning: Analysis of Spurious CoT Termination
arXiv:2609.03633v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning improves large reasoning models (LRMs) on complex tasks but often…
AI News Brief Hourly Summary 2026-09-04 18h : 12 posts
12 posts published in the last hour 15:33Test-time adaptation for speech enhancement with an autoregressive speech prior 15:33EraseSAE: Surgical Concept Erasure in Text-to-Video Diffusion Models via Sparse Autoencoders 15:33Remember and Reweight: Enhancing Multi-Agent Debate with Experience Memory and Confidence Estimation…
Test-time adaptation for speech enhancement with an autoregressive speech prior
arXiv:2609.03622v1 Announce Type: cross Abstract: Test-time adaptation (TTA) offers a promising direction for improving speech enhancement models under…
EraseSAE: Surgical Concept Erasure in Text-to-Video Diffusion Models via Sparse Autoencoders
arXiv:2609.03629v1 Announce Type: cross Abstract: Recent advances in text-to-video (T2V) diffusion models have demonstrated remarkable generative…
Remember and Reweight: Enhancing Multi-Agent Debate with Experience Memory and Confidence Estimation
arXiv:2609.03619v1 Announce Type: cross Abstract: Multi-agent debate (MAD) improves the reasoning capabilities of large language models by having multiple…
FailBench: How Reliable are VLMs at Judging Robot Task Success?
arXiv:2609.03611v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly used to evaluate robot manipulation outcomes, but…
ToolDF: Tool-Integrated Reasoning for Mixed-Authenticity Audio Deepfake Detection
arXiv:2609.03620v1 Announce Type: cross Abstract: Audio deepfake detection is commonly formulated as clip-level binary classification of single-domain…
LevelSyn: Physical-Aware Logic Synthesis via Level-Asynchronous Graph Neural Networks
arXiv:2609.03594v1 Announce Type: cross Abstract: As integrated circuit technology scales into the nanometer regime, the traditional disconnect between…
