Current models
Official catalog- GPT-6 Astra
- GPT-6.1 Sol
- GPT-6 Luna
- GPT-Image-2.5 Sunburst
- GPT-Image-2.5 Flare
- GPT-Live 1
- GPT-Realtime-2.1
- GPT-Realtime-2.1 Mini
- GPT-Realtime-2
- GPT-Realtime-Translate
- GPT-Realtime-1.5
- GPT-4o Mini TTS
- GPT-Transcribe
- GPT-Live-Transcribe
- GPT-Realtime-Whisper
- GPT-4o Transcribe
- GPT-4o Mini Transcribe
- text-embedding-3-small
- text-embedding-3-large
- GPT-5.6 Cyber / Daybreak RedRestricted access
- GPT-RosalindRestricted access
GPT-6.1 Sol
OpenAI updates Sol for coding, document understanding and computer use.
- Available through ChatGPT Work, Codex and the API; not yet in Chat.
Access at announcement
At launch, ChatGPT Work and Codex access covered Plus, Pro, Business, Enterprise and Edu. Official evaluations depend on the task and reasoning setting.
Announcements & evaluations
13 sourcesOpenAI lists Plus, Pro, Business, Enterprise and Edu access in ChatGPT Work and Codex at launch.
Current leaderboards
Intelligence Index v4.3.2. This is a checked snapshot; older method versions are not directly comparable.
View scoresUpdate date not stated
Update date not stated
| Claude Opus 5.5max with fallback | 58 |
|---|---|
| Claude Sonnet 5.5max with fallback | 56 |
| GPT-6 Astramax | 53 |
| Gemini 4 Argonhigh | 53 |
| GPT-6.1 Solmax | 52 |
| Qwen3.8 Max (0902) | 45 |
| Muse Spark 1.3max | 48 |
| GLM-5.3max | 45 |
| GLM-5.3low | 34 |
| GLM-5.3-Flash | 42 |
Text preference scores. Preliminary entries and uncertainty are retained; compare within this board.
View scoresOct 2, 2026
Data updated ·
| Gemini 4 Argonhigh; preliminary | 1525±9 |
|---|---|
| Claude Opus 5.5high | 1504±9 |
| Claude Fable 5.1max | 1501±6 |
| Gemini 3.8 Flashhigh; preliminary | 1495±5 |
| Muse Spark 1.3max | 1494±6 |
WebDev preference scores. Preliminary entries and uncertainty are retained; compare within this board.
View scoresOct 1, 2026
Data updated ·
| Claude Opus 5.5max | 1815+16/-16 |
|---|---|
| Claude Sonnet 5.5xhigh | 1786+18/-18 |
| GPT-6.1 Solmax | 1758+17/-17 |
| Gemini 4 Argonhigh; preliminary | 1680+13/-13 |
| Qwen3.8 Max (0902)preliminary | 1670+8/-8 |
Agent Arena measures Net Improvement, not task success rate. Scores depend on the listed configuration.
View scoresOct 2, 2026
Data updated ·
| Claude Fable 5.1max | 14.31%±1.90% |
|---|---|
| Claude Opus 5.5high | 13.82%±2.17% |
| Claude Sonnet 5.5max | 12.52%±3.09% |
| GPT-6 Astramax | 12.27%±2.23% |
| GPT-6.1 Solmax | 11.23%±2.76% |
TH1.1 task-completion time horizons. Model-specific chart values were not read; no scores are copied.
Data updated ·
Text-to-image preference scores; preliminary entries and generation settings are retained.
View scoresSep 24, 2026
Data updated ·
| gpt-image-2.5-sunburstPreliminary | 1424 ±8 |
|---|---|
| gpt-image-2.5-flarePreliminary | 1401 ±8 |
| muse-image | 1276 ±5 |
| gemini-3.1-flash-imagenano-banana-2; web-search | 1261 ±4 |
| seedream-5.0-pro | 1256 ±4 |
| qwen-image-3.0-pro | 1256 ±6 |
| qwen-image-2.1Preliminary | 1228 ±9 |
| flux-2-max | 1162 ±3 |
Blind listening preference with provider voices; not a controlled-voice or latency comparison.
View scoresUpdate date not stated
Update date not stated
| Eleven v495% CI1303–1339;1930samples;8nativevoices | 1321 ±18 |
|---|---|
| Qwen-Audio-3.1-TTS-Plus95% CI1274–1310;1481samples;8nativevoices | 1292 ±18 |
| Gemini3.8 Flash TTS95% CI1259–1291;2446samples;8nativevoices | 1275 ±16 |
| MiniMax Speech2.8 HD95% CI1162–1184;4638samples;8nativevoices | 1173 ±11 |
Family-wide source4
These sources cover the model family or historical versions. They do not evaluate this release.
Historical GPT-5.5 evaluation across five reasoning settings. This is not a test of GPT-6.1 Sol; results depend on the benchmarks and settings.
Historical software-task and selected-risk evaluation of gpt-5-thinking. Pre-release NDA access and publication approval are disclosed; results do not cover GPT-6.