14 posts published in the last hour
- 00:32LM Fight Arena: Benchmarking Large Multimodal Models via Game Competition
- 00:32MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use
- 00:32Bridging the Gap in Ophthalmic AI: MM-Retinal-Reason Dataset and OphthaReason Model toward Dynamic Multimodal Reasoning
- 00:32SurgRAW: Multi-Agent Workflow with Chain of Thought Reasoning for Robotic Surgical Video Analysis
- 00:32Rethinking Robot Safety in the Age of AI
- 00:32Enhancing knowledge tracing robustness for new question cold start in Intelligent Tutoring Systems
- 00:03Objective vs. Search: Decomposing What Makes a Good Tokeniser
- 00:03A Survey on Bridging EEG Signals and Generative AI: From Image and Text to Beyond
- 00:03A Zeroth-Order Paradigm for LLM Preference Alignment
- 00:03How workers are unlocking new ways of working
- 00:03Affora: A Design System for Agent-Friendly Interfaces
- 00:03How Cooley is accelerating IPO work with ChatGPT
- 00:03Dreaming the Sound of Contact: Leveraging Video and Audio Generation for Zero-Shot Force-Aware Manipulation and Data Generation
- 00:00AI News Brief Hourly Summary 2026-09-18 02h : 20 posts
