arXiv:2407.02765v4 Announce Type: replace-cross Abstract: We study the distributed optimization problem over a graphon with a continuum of nodes, which is…
On the Within-class Variation Issue in Alzheimer’s Disease Detection
arXiv:2409.16322v4 Announce Type: replace-cross Abstract: Alzheimer’s Disease (AD) detection commonly employs machine learning classification models to…
Learning When to Think: Adaptive Reasoning for Test-Time Compute Allocation
arXiv:2608.20256v2 Announce Type: replace Abstract: Reasoning language models trained with reinforcement learning typically operate under a fixed token…
AI-driven Prices for Externalities and Sustainability in Production Markets
arXiv:2106.06060v4 Announce Type: replace-cross Abstract: Traditional competitive markets do not account for negative externalities; indirect costs that…
Can We Trust AI Agents? A Case Study of an LLM-Based Multi-Agent System for Ethical AI
arXiv:2411.08881v3 Announce Type: replace-cross Abstract: AI-based systems, including Large Language Models (LLMs), impact millions by supporting diverse…
KernelArc: A Multi-Agent Framework for GPU Kernel Optimization
arXiv:2608.17071v2 Announce Type: replace Abstract: We present KernelArc, a multi-agent framework for autonomous GPU kernel optimization across…
DECOWAM: Decoupled Whole-Body World-Action Model for Legged Mobile Manipulation
arXiv:2608.20114v2 Announce Type: replace Abstract: Mobile manipulation requires a robot to predict how locomotion and arm motion jointly alter future…
FM-Bench: A Benchmark for Long-Horizon Management with Competing Agents
arXiv:2608.18423v2 Announce Type: replace Abstract: Language model agents now execute bounded tasks reliably. Whether they can sustain effective…
The Lifecycle of LLM-as-a-Judge for Large-Scale Recommendation Explanations
arXiv:2608.18300v2 Announce Type: replace Abstract: LLM-as-a-Judge, which leverages a large language model to evaluate natural language generated by…
Attributing Preprocessing Invariance in Spectral Foundation Models
arXiv:2608.14227v2 Announce Type: replace Abstract: Preprocessing invariance is an appealing goal for spectral foundation models: a frozen model should…
AI News Brief Hourly Summary 2026-08-25 04h : 11 posts
11 posts published in the last hour 01:32Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces 01:32TRACE-Memory: Public-Conditioned Retrieval and Utility-Aware Evidence Admission for Personalized Generation 01:32MemWM: Memory-Augmented Text-Based World Model 01:31Towards Query-Agnostic RAG…
Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces
arXiv:2608.11354v2 Announce Type: replace Abstract: Modern recommender systems treat observed actions as reliable proxies for user preferences, yet…
TRACE-Memory: Public-Conditioned Retrieval and Utility-Aware Evidence Admission for Personalized Generation
arXiv:2608.08446v2 Announce Type: replace Abstract: Personalized generation systems retrieve user history by request–memory relevance and inject it into…
MemWM: Memory-Augmented Text-Based World Model
arXiv:2608.07107v2 Announce Type: replace Abstract: World models are increasingly used to support planning in agents by predicting how environment states…
Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability
arXiv:2608.11238v2 Announce Type: replace Abstract: Retrieval-augmented generation improves the factuality of large language models by grounding responses…
Fragility of Value under Imperfect Alignment
arXiv:2607.28881v4 Announce Type: replace Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that…
The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs
arXiv:2607.08734v2 Announce Type: replace Abstract: Post-Training Quantization has become widely used to compress large language models to make them…
Share the Judge, Learn the Deferral: Where Specialization Helps LLM Evaluation
arXiv:2607.27984v2 Announce Type: replace Abstract: Agentic systems generate outputs faster than human review. We contrast two LLM evaluator…
