Baseten Adds DeepSeek-V4.1-Flash to Model APIs With 1M-Token Context

DeepSeek-V4.1-Flash is available now on Baseten Model APIs, Baseten announced on September 11, 2026, bringing the 552B-parameter multimodal mixture-of-experts (MoE) model, which pairs 8B active parameters for prefill with 16B for decode across a 1M-token context window, to the inference provider's platform. DeepSeek released the model's open weights on Hugging Face, and DeepSeek's own announcement is dated September 9, 2026. The model accepts text and image input and generates text output, and…

This article has been indexed from Unite.AI

Read the original article: