AI models

DeepSeek V

Open timeline

Current models

Official catalog
  • DeepSeek-V4.1-Flash
  • DeepSeek-V4-Pro-0813
Release recordReleased

DeepSeek V4.1 Flash

V4.1 Flash introduces native visual understanding and an asymmetric MoE architecture.

  • The API identifier is deepseek-flash.

Access at announcement

Legacy V4 API names can route to Flash; a name alone does not identify the served version.

Full release

Announcements & evaluations

9 sources
Official announcementsOfficial
DeepSeek V4.1 Flash:更强、更快、更普惠
Evidence & scope

Original checked · Oct 5, 2026

Read original

DeepSeek V4.1 Flash

Model cards & docsOfficial
DeepSeek-V4.1-Flash · Hugging Face
Evidence & scope

Original checked · Oct 5, 2026

Read original

DeepSeek V4.1 Flash

Social postsOfficial
DeepSeek V4.1-Flash API migration

Official API post identifies deepseek-flash and the temporary routing of older V4 model names. Performance comparisons are maker claims.

Evidence & scope

Original checked · Oct 5, 2026

Read original

DeepSeek V4.1 Flash

Current leaderboards

Independent evaluations
Artificial Analysis LLM Leaderboard · 4.3.2

Intelligence Index v4.3.2. This is a checked snapshot; older method versions are not directly comparable.

View scoresUpdate date not stated

Update date not stated

Artificial Analysis Intelligence Index · Checked snapshot Oct 5, 2026
Claude Opus 5.5max with fallback58
Claude Sonnet 5.5max with fallback56
GPT-6 Astramax53
Gemini 4 Argonhigh53
GPT-6.1 Solmax52
Qwen3.8 Max (0902)45
Muse Spark 1.3max48
GLM-5.3max45
GLM-5.3low34
GLM-5.3-Flash42
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Text Arena Overall

Text preference scores. Preliminary entries and uncertainty are retained; compare within this board.

View scoresOct 2, 2026

Data updated ·

Arena score · Checked snapshot Oct 5, 2026
Gemini 4 Argonhigh; preliminary1525±9
Claude Opus 5.5high1504±9
Claude Fable 5.1max1501±6
Gemini 3.8 Flashhigh; preliminary1495±5
Muse Spark 1.3max1494±6
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Code Arena WebDev Overall

WebDev preference scores. Preliminary entries and uncertainty are retained; compare within this board.

View scoresOct 1, 2026

Data updated ·

Arena score · Checked snapshot Oct 5, 2026
Claude Opus 5.5max1815+16/-16
Claude Sonnet 5.5xhigh1786+18/-18
GPT-6.1 Solmax1758+17/-17
Gemini 4 Argonhigh; preliminary1680+13/-13
Qwen3.8 Max (0902)preliminary1670+8/-8
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Agent Arena Overall

Agent Arena measures Net Improvement, not task success rate. Scores depend on the listed configuration.

View scoresOct 2, 2026

Data updated ·

Net Improvement · Checked snapshot Oct 5, 2026
Claude Fable 5.1max14.31%±1.90%
Claude Opus 5.5high13.82%±2.17%
Claude Sonnet 5.5max12.52%±3.09%
GPT-6 Astramax12.27%±2.23%
GPT-6.1 Solmax11.23%±2.76%
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Task-Completion Time Horizons of Frontier AI Models · TH1.1

TH1.1 task-completion time horizons. Model-specific chart values were not read; no scores are copied.

Data updated ·

Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Family-wide source1

These sources cover the model family or historical versions. They do not evaluate this release.

Model cards & docsOfficialFamily-wide source
Models & Pricing | DeepSeek API Docs
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source