ModelRadar
/

DeepSeek V4.1 Flash

DeepSeek🇨🇳 CNFrontier· Sep 10, 2026

TextReasoningVisionopen weightsAPI

Multimodal MoE model that natively processes images and text with a 1M-token context window. Based on a 552B backbone, it activates 8B/16B per step, cutting KV-cache overhead to roughly a quarter of its predecessor.

Summary of the model card

Context
1M tokens
Price / 1M tok
$0.04 in · $0.29 out
License
mit
Modalities
text, image → text

3 variants

Standard1M$0.04/$0.29OR
Batch1M$0.11/$0.34OR
StandardHF

Sources

HuggingFaceOpenRouter
Deutsch

Multimodales MoE-Modell, das Bilder und Text nativ verarbeitet und bis zu eine Million Tokens Kontext unterstützt. Es basiert auf einem 552B-Backbone, aktiviert pro Schritt nur 8B bzw. 16B Parameter und reduziert den KV-Cache-Aufwand auf etwa ein Viertel der Vorgängerversion.

Benchmarks

BenchmarkDomainValueEvidence
LongBench v2vendor numberlong-context44.7 % correct90
SimpleQAvendor numberknowledge30.1 % correct81
MMMU-Provendor numbervision56.5 % correct80
DocVQAvendor numbervision95.6 ANLS71
BigCodeBenchvendor numbercoding56.8 % pass@166
HumanEvalvendor numbercoding69.5 % pass@157
GPQA Diamondvendor numberreasoning93.4 % correct50
MGSMvendor numbermultilingual84.4 % correct44
MMLUvendor numberknowledge68.3 % correct43

Vendor-reported numbers are marked and count with factor 0.7 in rankings. The evidence score says how much a benchmark still tells you today.

Provenance

Detected
Sep 27, 2026
First source
openrouter
Description
Summary of the model card
Editorially reviewed
—