DeepSeek

DeepSeek V4.1 FlashAberto

DeepSeek V4.1 Flash is an open-weights MoE reasoning model with 552B total (16B active) parameters, 1.0M context, and chain-of-thought problem solving.

#317lançado 10 de set. de 2026·Open Weights·552B (16B active)
textimage1.0M ctxreasoning
API providersRanking
Intelligence
54.5 #317
AAQI 40.0
Output speed
214.4 t/s
TTFT 1.21 s · ITL 4.7 ms/tok
Blended price (3:1)
$0.18 /1M
In $0.3/M · Out $1.2/M
Context
1.0M
≈ 2,500 A4 pages · max out 32768

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

DeepSeek V4.1 Flash
Claude Mythos Preview (#1)
Average (68)
Coding (52)Mathematics (68)Hard Science (GPQA) (58)Agentic & Tool Use (48)Knowledge (MMLU) (46)Vision / Multimodal (50)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see Metodologia.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

Pontuações por fonte

Detalhes do modelo

Data de lançamento
10 de set. de 2026
Parâmetros
552B (16B active)
Licença
Open Weights
Janela de contexto
1,000,000 tokens
Preço de entrada
$0.3/M
Preço de saída
$1.2/M
Velocidade de saída
214 t/s
Tempo até o primeiro token
1.21 s

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/deepseek-v4-1-flash/rank.svg)](https://llmpodium.com/models/deepseek-v4-1-flash)
BibTeX
@misc{llmpodium2026deepseekv41flash,
  title = {Benchmark Analysis and Intelligence Evaluation of DeepSeek V4.1 Flash},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/deepseek-v4-1-flash}
}

More from DeepSeek

Modelos alternativos

Principais alternativas para DeepSeek V4.1 Flash com base em recursos, preço e benchmarks.

0 / 4