11 posts were published in the last hour
- 8:32 : Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution
- 8:32 : REIN: Bridging the Gap between Reasoning and Reliability via Reflection and Abstention Alignment
- 8:32 : TongGuOCR: A Layout-Aware and Token-Augmented OCR Framework for Chinese Historical Documents
- 8:32 : When Is Benchmark Contamination Detectable? Information Limits and Power-Calibrated Audits
- 8:32 : ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration
- 8:3 : Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills
- 8:3 : TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?
- 8:3 : GRACE: LLM-Grounded Semantic Metric Spaces for Scalable Mixed-Data Clustering
- 8:2 : SurgLAT: Surgical Latent Attention Tracking for Depth-Aware Robotic Laparoscope Control
- 8:2 : GraphThink: Graph-Enhanced LLM Thinking for Long-Horizon Embodied Task Planning
- 8:0 : AI News Brief Hourly Summary 2026-08-11 10h : 12 posts