DeepSeek's 10 Sep 2026 Flash refresh: 552B MoE, 8B active on prefill and 16B on decode, native vision, 1M context, 384K max output. API id deepseek-flash. Official off-peak list is $0.15 / $0.60 per 1M tokens (cache-miss input / output); peak doubles. Retired V4-Flash aliases route here.
Current DeepSeek Flash API model for cheap multimodal and agent work. Prefer official pricing and architecture notes over vendor eval tables; AA rows land when the snapshot lists the slug.
Not enough families covered yet.Add more benchmark rows for DeepSeek V4.1 Flash to see the family profile.
Editor's note
Recorded from DeepSeek's 10 Sep 2026 launch post, API pricing table, and Hugging Face model card (552B MoE, CED 8B/16B active, 1M context, MIT weights). Vendor Terminal-Bench / DeepSWE / AutomationBench figures stay off the ledger. No independent AA Intelligence Index or Vals Index row as of 12 Sep 2026.
Benchmark placements
Per-benchmark positions
Where DeepSeek V4.1 Flash places on each public benchmark source that publishes a row. Rank counts every model with a latest row in the same test — not a universal quality score.
No benchmark rows yet for DeepSeek V4.1 Flash.
Family context
Strongest rows in each family
The three highest-scoring models with pages in each capability family. Where DeepSeek V4.1 Flash shows up, it's highlighted.