arXiv:2609.29245v1 Announce Type: cross Abstract: Given a large corpus, the questions one might ask can vary — from “When was the first human heart…
Author: script
Spot, Separate, and Enhance: Fully Generative Approach for Audio Mixing
arXiv:2609.29169v1 Announce Type: cross Abstract: We introduce Spot, Separate, and Enhance (SSE), the first multimodal, user-guided generative model for…
Not Every Token Is Worth Distilling: Selective Supervision for Direct-OPD
arXiv:2609.29142v1 Announce Type: cross Abstract: Direct On-Policy Distillation (Direct-OPD) transfers reinforcement-learning-induced policy improvements…
AI-Moderated Interviews for Market Research and Digital Twins Calibration
arXiv:2609.29143v1 Announce Type: cross Abstract: AI-moderated interviews are emerging as a scalable market-research method for generating consumer…
Med-AR: Autoregressive Vision-Language Pretraining for Long-Tailed Chest X-Ray Classification and Uncertainty-Aware Evaluation
arXiv:2609.29156v1 Announce Type: cross Abstract: Long-tailed chest X-ray classification requires visual representations that capture both common…
HarnessPAI: An Evolving Harness for Physical AI
arXiv:2609.29166v1 Announce Type: cross Abstract: Physical AI aims to build embodied agents that perceive the world, understand and reason about it, and…
AI News Brief Hourly Summary 2026-09-25 22h : 15 posts
15 posts published in the last hour 19:33DAWN: Noise-Robust Quadruped Parkour via Depth-Denoising World Models 19:32Tag-Aware Structured Text Translation: Towards a Systematic Understanding 19:32WildHSR: Metric Feed-Forward 4D People-Scene Reconstruction from a 3D Foundation Model 19:32Where Does Exactly-Once Live? Model, Harness,…
DAWN: Noise-Robust Quadruped Parkour via Depth-Denoising World Models
arXiv:2609.29092v1 Announce Type: cross Abstract: Vision-based legged locomotion methods assume clean depth at training time and rely on hand-tuned…
Tag-Aware Structured Text Translation: Towards a Systematic Understanding
arXiv:2609.29131v1 Announce Type: cross Abstract: Internet texts are replete with format tags that carry structural, semantic, and functional meaning.…
WildHSR: Metric Feed-Forward 4D People-Scene Reconstruction from a 3D Foundation Model
arXiv:2609.29106v1 Announce Type: cross Abstract: 3D foundation models recover video cameras and geometry in one forward pass, but some of the strongest…
