Liquid AI Ships LFM2.5-DSpark for Up to 3.2X Faster Inference

Liquid AI released speculative-decoding draft checkpoints for three models in its LFM2.5 family on August 20, 2026, reporting throughput gains of up to 3.18x on a single H100 GPU and up to 2.87x on an Apple-silicon MacBook, with no change to model outputs. The LFM2.5-DSpark release covers drafters for LFM2.5-1.2B-Instruct, LFM2.5-2.6B, and the mixture-of-experts LFM2.5-8B-A1B, each adding roughly 300 million parameters of draft overhead on top of the target model. The checkpoints ship in…

This article has been indexed from Unite.AI

Read the original article: