AI models

Mistral models

Open timeline

Current models

Official catalog
  • Mistral Medium 3.5
  • Mistral Small 4
  • Mistral Large 3
  • Ministral 3 14B
  • Ministral 3 8B
  • Ministral 3 3B
  • OCR 4.1
  • Voxtral TTS
  • Voxtral Mini Transcribe 2
  • Voxtral Mini Transcribe Realtime
  • Voxtral Small
  • Codestral
  • Codestral Embed
  • Mistral Embed
  • Shieldstral 1.0
  • Mistral Moderation 2
Release recordBeta / early access

Mistral Medium 3.5 announcement

Mistral announces a 128B model with visual input and adjustable reasoning effort.

  • Open weights use a Modified MIT license.

Access at announcement

The May 22 news article describes public preview. The model directory separately dates the API variant April 28 and marks it GA.

Full release

Announcements & evaluations

9 sources
Official announcementsOfficial
Remote agents in Vibe. Powered by Mistral Medium 3.5.
Evidence & scope

Original checked · Oct 5, 2026

Read original

Mistral Medium 3.5 announcement

Independent evaluations
Factuality in the Arena

Samples arena battles and checks web-verifiable factual claims, combining factuality with human preference. Preference is not a pure factual-accuracy score.

Evidence & scope

Original checked · Oct 5, 2026

Read original

Mistral Medium 3.5 announcement

Current leaderboards

Independent evaluations
Artificial Analysis LLM Leaderboard · 4.3.2

Intelligence Index v4.3.2. This is a checked snapshot; older method versions are not directly comparable.

View scoresUpdate date not stated

Update date not stated

Artificial Analysis Intelligence Index · Checked snapshot Oct 5, 2026
Claude Opus 5.5max with fallback58
Claude Sonnet 5.5max with fallback56
GPT-6 Astramax53
Gemini 4 Argonhigh53
GPT-6.1 Solmax52
Qwen3.8 Max (0902)45
Muse Spark 1.3max48
GLM-5.3max45
GLM-5.3low34
GLM-5.3-Flash42
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Text Arena Overall

Text preference scores. Preliminary entries and uncertainty are retained; compare within this board.

View scoresOct 2, 2026

Data updated ·

Arena score · Checked snapshot Oct 5, 2026
Gemini 4 Argonhigh; preliminary1525±9
Claude Opus 5.5high1504±9
Claude Fable 5.1max1501±6
Gemini 3.8 Flashhigh; preliminary1495±5
Muse Spark 1.3max1494±6
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Code Arena WebDev Overall

WebDev preference scores. Preliminary entries and uncertainty are retained; compare within this board.

View scoresOct 1, 2026

Data updated ·

Arena score · Checked snapshot Oct 5, 2026
Claude Opus 5.5max1815+16/-16
Claude Sonnet 5.5xhigh1786+18/-18
GPT-6.1 Solmax1758+17/-17
Gemini 4 Argonhigh; preliminary1680+13/-13
Qwen3.8 Max (0902)preliminary1670+8/-8
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Agent Arena Overall

Agent Arena measures Net Improvement, not task success rate. Scores depend on the listed configuration.

View scoresOct 2, 2026

Data updated ·

Net Improvement · Checked snapshot Oct 5, 2026
Claude Fable 5.1max14.31%±1.90%
Claude Opus 5.5high13.82%±2.17%
Claude Sonnet 5.5max12.52%±3.09%
GPT-6 Astramax12.27%±2.23%
GPT-6.1 Solmax11.23%±2.76%
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Task-Completion Time Horizons of Frontier AI Models · TH1.1

TH1.1 task-completion time horizons. Model-specific chart values were not read; no scores are copied.

Data updated ·

Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Independent evaluations
Artificial Analysis TTS Provider Voice · Provider Voice

Blind listening preference with provider voices; not a controlled-voice or latency comparison.

View scoresUpdate date not stated

Update date not stated

Provider Voice Elo · Checked snapshot Oct 5, 2026
Eleven v495% CI1303–1339;1930samples;8nativevoices1321 ±18
Qwen-Audio-3.1-TTS-Plus95% CI1274–1310;1481samples;8nativevoices1292 ±18
Gemini3.8 Flash TTS95% CI1259–1291;2446samples;8nativevoices1275 ±16
MiniMax Speech2.8 HD95% CI1162–1184;4638samples;8nativevoices1173 ±11
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source

Family-wide source1

These sources cover the model family or historical versions. They do not evaluate this release.

Model cards & docsOfficialFamily-wide source
Models Overview | Mistral Docs
Evidence & scope

Original checked · Oct 5, 2026

Read original

Family-wide source