DeepSeek-V4.1-Flash is available now on Baseten Model APIs, Baseten announced on September 11, 2026, bringing the 552B-parameter multimodal mixture-of-experts (MoE) model, which pairs 8B active parameters for prefill with 16B for decode across a 1M-token context window, to the inference provider's platform. DeepSeek released the model's open weights on Hugging Face, and DeepSeek's own announcement is dated September 9, 2026. The model accepts text and image input and generates text output, and…
Read the original article:
