DeepSeekV4.1 Flash

External Link

DeepSeek-V4.1-Flash is DeepSeek's open-weight multimodal mixture-of-experts model with a 552B-parameter backbone, phase-dependent 8B/16B activation, one-million-token context, and continuously adjustable reasoning effort.

Token context
1,049K tokens
Inputs
Text, Image
Outputs
Text
Released
Sep 10, 2026

Benchmark Market Position

Intelligence, Cost, and Efficiency

All rankings →

The orange outline marks DeepSeek V4.1 Flash. Each available panel preserves its real rank and value; missing benchmark or pricing inputs remain explicitly unranked.

Providers & Pricing

Estimate workload cost →
ProviderInput / 1MOutput / 1MCache read / 1MCache write / 1MObservedEvidence
$0.20$0.60$0.006Sep 22, 2026Provider-reported
$0.15$0.60$0.003Sep 10, 2026Official fact
LiveBench2026-06-2583.20percent36,3551,270deepseek-v4.1-flash-maxaverage per caseRecomputedwinner eligibleThird-party benchmark
LMArena Agent Arenaagent-2026-09-15-d25aabda00104.88score20,080Deepseek V4.1 Flash (Max); 95% CI [3.54852933, 6.21109128]; sessions 20080; observations 2295632; rank 12unknownOriginal sourcewinner eligibleThird-party benchmark
ToneBench2026-09-11-10-task-4ef099199c9c88.03points18,19950DeepSeek V4.1 Flash (max)average per caseOriginal sourcenot winner eligibleThird-party benchmark
ToneBench2026-09-11-10-task-4ef099199c9c87.13points10,01150DeepSeek V4.1 Flash (default)average per caseOriginal sourcenot winner eligibleThird-party benchmark

At a Glance

Model Facts

Identity

Status
active
Developer ID
deepseek-ai/DeepSeek-V4.1-Flash
Released
Sep 10, 2026
Version
DeepSeek-V4.1-Flash
License
mit

Capacity

Token context
1,049K tokens
Maximum output
393K tokens
Total parameters
763.2B
Active parameters
Unknown
Knowledge cutoff
Not reported

Interface and Access

Inputs
Text, Image
Outputs
Text
API available
Yes
Open weights
Yes
Self-hostable
Yes
Reasoning
Yes
Vision
Yes
Tool calling
Yes
Capabilities
agents, chat, fim, generation, reasoning, responses, structured outputs, tools, vision

Technical Specifications

Architecture design
Causal Encoder-Decoder (20 encoder + 20 decoder layers)
Backbone parameters
552,000,000,000 parameters
Active parameters during prefill
8,000,000,000 parameters
Active parameters during decode
16,000,000,000 parameters
Transformer layers
40 layers
Routed experts per MoE layer
384 experts
Routed experts per token
6 experts
Reasoning effort range
1–100
Pre-training corpus
45,000,000,000,000 tokens

This page represents one developer model product. Dated API snapshots, serving endpoints, pricing, regions, quantizations, and service tiers attach as versioned aliases or provider details and never create duplicate public model records.

Version Lineage

PredecessorUnknown
SuccessorsUnknown
Aliasesdeepseek-ai/DeepSeek-V4.1-Flash, deepseek-flash, DeepSeek V4.1, deepseek-v4-flash, deepseek-v4-flash-vision-exp

Explicit Unknowns

  • Knowledge cutoff

Primary Evidence

Sources and Observation Date

Comparable Peers

Related Model Comparisons

All comparisons →
APairBContext
vsfamily variantstext
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peerstext
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peersimage, text

Questions

DeepSeek V4.1 Flash FAQs

What is DeepSeek V4.1 Flash?+

DeepSeek-V4.1-Flash is DeepSeek's open-weight multimodal mixture-of-experts model with a 552B-parameter backbone, phase-dependent 8B/16B activation, one-million-token context, and continuously adjustable reasoning effort. It is developed by DeepSeek and its current sourced lifecycle status is active.

What inputs and outputs does DeepSeek V4.1 Flash support?+

DeepSeek V4.1 Flash's current record lists Text and Image as input and Text as output.

How much context does DeepSeek V4.1 Flash support?+

The current record lists a 1,049K-token context window and a maximum output of 393K tokens.

What technical specifications are published for DeepSeek V4.1 Flash?+

DeepSeek V4.1 Flash's developer-published specifications include architecture design: Causal Encoder-Decoder (20 encoder + 20 decoder layers); backbone parameters: 552000000000 parameters; active parameters during prefill: 8000000000 parameters; active parameters during decode: 16000000000 parameters; transformer layers: 40 layers; routed experts per moe layer: 384 experts; routed experts per token: 6 experts; reasoning effort range: 1–100; pre-training corpus: 45000000000000 tokens.

How large is DeepSeek V4.1 Flash?+

The current record lists 763.2B total parameters and an unknown active parameter count. Its architecture is DeepseekV41ForCausalLM.

What capabilities and tasks does DeepSeek V4.1 Flash support?+

DeepSeek V4.1 Flash's recorded capabilities are agents, chat, fim, generation, reasoning, responses, structured outputs, tools, and vision. Its supported tasks are image, reasoning, and text.

Is DeepSeek V4.1 Flash available through an API?+

DeepSeek V4.1 Flash is marked as API-available. The directly sourced provider records currently include Deepinfra, DeepSeek, and Together Ai.

How much does DeepSeek V4.1 Flash cost through an API?+

The lowest directly sourced prices currently attached to DeepSeek V4.1 Flash are $0.15 per million input tokens through DeepSeek and $0.60 per million output tokens through Deepinfra. Prices are provider-specific and should be checked against each cited observation date.

Can DeepSeek V4.1 Flash be self-hosted?+

DeepSeek V4.1 Flash is an open-weight model. Self-hosting is supported, and the recorded license is mit.

Does DeepSeek V4.1 Flash have related versions or aliases?+

No predecessor is recorded. No successor is recorded. Recorded aliases are deepseek-ai/DeepSeek-V4.1-Flash, deepseek-flash, DeepSeek V4.1, deepseek-v4-flash, and deepseek-v4-flash-vision-exp.

Send Feedback