18 posts were published in the last hour 17:33 : Learning from Multimodal Pseudo-Labels for Robust Open-Vocabulary Instance and Panoptic Segmentation 17:33 : Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets 17:32…
Learning from Multimodal Pseudo-Labels for Robust Open-Vocabulary Instance and Panoptic Segmentation
arXiv:2608.11681v1 Announce Type: cross Abstract: This work addresses the challenge of open-vocabulary instance segmentation (OVIS) and open-set panoptic…
Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets
This post has no text preview — click the link below to read the original article. This article has been indexed from Hugging Face – Blog Read the original article: Record, train, and deploy from one place with Strands Agents,…
The Wording Effect: Quantifying Two-Way Drift in LLM Benchmark Performance
arXiv:2608.11694v1 Announce Type: cross Abstract: A benchmark score comes from a single phrasing of each problem. That single phrasing is treated as if it…
OpenAI hires new CRO as executive shake-up continues
Dali Rajic will take over as OpenAI’s top salesperson.
APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference
arXiv:2608.11688v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models are attractive for edge deployment because they provide high model…
Introducing Gemini 3.7 Flash
This post has no text preview — click the link below to read the original article. This article has been indexed from Google DeepMind News Read the original article: Introducing Gemini 3.7 Flash
Consolidator: Learning Persistent Routed Memory Across Context Boundaries
arXiv:2608.11701v1 Announce Type: cross Abstract: Copying short-term memory (STM) into a slower store can preserve state across a context boundary, but…
Bring your spreadsheet data to life with Sheets canvas
The video shows Sheets canvas in action.
REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation
arXiv:2608.11698v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories under dense token-level…
Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperative Multi-Agent Reinforcement Learning
arXiv:2608.11658v1 Announce Type: cross Abstract: Many reinforcement learning systems, from fleet management to traffic signal control, must serve an…
GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs
arXiv:2608.11674v1 Announce Type: cross Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they…
Amazon Quick Arrives Inside Word, Excel, PowerPoint, and Outlook
AWS has brought its Amazon Quick assistant directly into Microsoft 365, announcing on August 13, 2026 that extensions for Word, Excel, PowerPoint, and…
Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing
arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are…
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster. Powered by Cerebras, it delivers up to 750 output tokens per…
Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL
arXiv:2608.11669v1 Announce Type: cross Abstract: Reinforcement learning against rubrics, lists of criteria graded by an LLM judge, has become a standard…
What We Learned by Reproducing 2,200 papers from ICML
This post has no text preview — click the link below to read the original article. This article has been indexed from Hugging Face – Blog Read the original article: What We Learned by Reproducing 2,200 papers from ICML
Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads
arXiv:2608.11661v1 Announce Type: cross Abstract: A multiplicative dual-encoder network computes a real-valued output for a pair of inputs as the inner…
