model fatıgue
Comparison, updated in place

The cheapest model at every Artificial Analysis score

Published 4 Oct 2026Data checked 4 Oct 2026Updated 4 Oct 2026Cite this readingEvery reading

The short answer

On Artificial Analysis's leaderboard as we read it on 4 October 2026, 102 current model settings have a score and a cost per task above zero. Only 15 of them are beaten by no other setting on score and cost at once: GPT-6 Luna at every reasoning setting from low to max, Xiaomi's MiMo-V2.6-Flash and MiMo-V2.6-Pro, GPT-6.1 Sol at every setting, and Claude Opus 5.5 at high, xhigh and max.

Score against cost per task, every current setting with a cost per task

Artificial Analysis Intelligence Index up, cost per task across on a log scale. The line is the best score on the list at or under each cost.

No other setting beats it on bothEvery other setting with a cost per task
0102030405060$0.01$0.1$1$10Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback): score 56.0, $7.67 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback): score 53.4, $7.63 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback): score 53.2, $5.98 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Max): score 52.7, $3.26 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 4 Argon (High): score 52.6, $1.99 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Xhigh): score 52.4, $2.31 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback): score 51.9, $2.75 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback): score 51.2, $1.34 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback): score 51.2, $3.91 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (High): score 50.9, $1.73 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Medium): score 49.6, $1.54 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback): score 48.9, $2.98 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMuse Spark 1.3 (Max): score 48.1, $1.60 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback): score 46.8, $2.37 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, High Effort, Default Fallback): score 46.8, $1.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGrok 4.7 (Xhigh): score 46.4, $3.74 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGrok 4.7 (High): score 46.3, $2.73 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Low): score 45.8, $0.82 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 Max (0902): score 45.4, $5.41 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMuse Spark 1.3 (Xhigh): score 45.1, $1.37 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGLM-5.3 (Max): score 44.8, $2.01 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTStep 5 Preview: score 43.7, $0.72 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTKimi K3 (Max): score 43.6, $2.00 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Low Effort, Default Fallback): score 42.3, $0.55 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGrok 4.7 (Low): score 42.2, $1.25 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Max): score 42.1, $1.40 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGLM 5.3 Flash: score 41.8, $0.25 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.8 Flash (High): score 40.9, $1.24 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback): score 40.8, $0.59 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 2.4T A95B: score 39.9, $2.16 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8-Flash-Next: score 39.8, $0.37 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.8 Flash (Medium): score 39.8, $0.93 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4.1 Flash (Max): score 39.5, $0.27 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Xhigh): score 38.0, $0.63 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4 Pro 0813 (Max): score 36.0, $0.67 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, Low Effort, Default Fallback): score 35.9, $0.42 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4 Flash Vision (Max): score 34.8, $0.31 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGLM-5.3 (Low): score 34.3, $0.85 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (High): score 34.2, $0.34 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Xhigh): score 33.7, $1.01 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Medium): score 30.1, $0.18 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTKimi K3 (Low): score 30.1, $1.15 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.1 Pro Preview: score 29.7, $0.67 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMiniMax-M3: score 29.2, $0.51 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Medium): score 27.6, $1.13 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Low): score 27.5, $0.14 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQuasar 438B (Max, Based on GLM-5.2): score 26.7, $2.02 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTApodex 1.1: score 26.4, $0.46 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Low): score 26.2, $1.05 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.5 Instant (June 2026): score 26.0, $0.69 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTKimi K2.7 Code: score 25.8, $0.54 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTInkling Small: score 25.7, $0.089 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTHy3: score 25.3, $0.072 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.7 Plus: score 25.2, $0.22 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4.1 Flash (Non-reasoning): score 24.7, $0.15 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTSolar Mini 4: score 24.1, $0.36 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNemotron 3 Ultra 550B A55B (Reasoning): score 22.9, $0.60 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.5 Flash-Lite: score 22.2, $0.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Non-reasoning): score 20.8, $0.14 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4 Pro 0813 (Non-reasoning): score 20.4, $0.48 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Non-reasoning): score 20.2, $2.49 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTLongCat 2.0: score 19.1, $0.059 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Non-reasoning): score 18.5, $0.011 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.5 397B A17B (Reasoning): score 18.4, $0.47 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.6 35B A3B (Reasoning): score 18.2, $0.48 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMuse Glimmer (High): score 17.5, $0.057 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude 4.5 Haiku (Reasoning): score 16.9, $0.28 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTRing-2.6-1T: score 16.6, $0.29 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.5 122B A10B (Reasoning): score 15.6, $0.32 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMistral Medium 3.5: score 14.2, $0.50 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNemotron 3.5 Lightning: score 12.9, $0.093 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNemotron 3 Super 120B A12B (Reasoning): score 12.8, $1.64 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMercury 2.5: score 12.3, $0.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTgpt-oss-120b (High): score 11.6, $0.11 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMistral Small 4 (Reasoning): score 11.3, $0.015 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.5 9B (Reasoning): score 11.2, $0.21 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGranite 4.2 8B: score 11.1, $0.024 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTTrinity Large Thinking: score 10.8, $0.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMistral Large 3: score 9.3, $0.031 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3 Coder Next: score 9.2, $0.55 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGranite 4.2 3B: score 9.1, $0.0060 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTgpt-oss-20b (High): score 9.0, $0.012 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNVIDIA Nemotron 3 Nano 30B A3B (Reasoning): score 8.9, $0.017 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTCeleris-1: score 6.3, $0.050 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMinistral 3 14B: score 6.0, $0.019 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMinistral 3 8B: score 5.5, $0.011 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMinistral 3 3B: score 4.8, $0.0078 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback): score 57.6, $5.98 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback): score 56.0, $3.46 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback): score 53.6, $1.82 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Max): score 51.8, $0.72 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Xhigh): score 51.0, $0.39 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (High): score 50.2, $0.32 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Medium): score 47.8, $0.21 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMiMo-V2.6-Pro: score 46.3, $0.13 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Low): score 42.1, $0.13 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Max): score 38.1, $0.068 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMiMo-V2.6-Flash: score 37.9, $0.062 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Xhigh): score 34.6, $0.042 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (High): score 32.9, $0.029 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Medium): score 29.9, $0.017 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Low): score 21.5, $0.0045 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 LunaMiMo-V2.6-ProGPT-6.1 SolOpus 5.5Sonnet 5.5 maxFable 5.1 maxGemini 4 Argonscorecost per task, log scale
0102030405060$0.01$0.1$1$10Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback): score 56.0, $7.67 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback): score 53.4, $7.63 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback): score 53.2, $5.98 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Max): score 52.7, $3.26 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 4 Argon (High): score 52.6, $1.99 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Xhigh): score 52.4, $2.31 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback): score 51.9, $2.75 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback): score 51.2, $1.34 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback): score 51.2, $3.91 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (High): score 50.9, $1.73 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Medium): score 49.6, $1.54 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback): score 48.9, $2.98 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMuse Spark 1.3 (Max): score 48.1, $1.60 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback): score 46.8, $2.37 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, High Effort, Default Fallback): score 46.8, $1.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGrok 4.7 (Xhigh): score 46.4, $3.74 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGrok 4.7 (High): score 46.3, $2.73 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Astra (Low): score 45.8, $0.82 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 Max (0902): score 45.4, $5.41 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMuse Spark 1.3 (Xhigh): score 45.1, $1.37 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGLM-5.3 (Max): score 44.8, $2.01 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTStep 5 Preview: score 43.7, $0.72 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTKimi K3 (Max): score 43.6, $2.00 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Low Effort, Default Fallback): score 42.3, $0.55 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGrok 4.7 (Low): score 42.2, $1.25 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Max): score 42.1, $1.40 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGLM 5.3 Flash: score 41.8, $0.25 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.8 Flash (High): score 40.9, $1.24 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback): score 40.8, $0.59 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 2.4T A95B: score 39.9, $2.16 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8-Flash-Next: score 39.8, $0.37 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.8 Flash (Medium): score 39.8, $0.93 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4.1 Flash (Max): score 39.5, $0.27 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Xhigh): score 38.0, $0.63 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4 Pro 0813 (Max): score 36.0, $0.67 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Sonnet 5.5 (Adaptive Reasoning, Low Effort, Default Fallback): score 35.9, $0.42 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4 Flash Vision (Max): score 34.8, $0.31 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGLM-5.3 (Low): score 34.3, $0.85 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (High): score 34.2, $0.34 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Xhigh): score 33.7, $1.01 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Medium): score 30.1, $0.18 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTKimi K3 (Low): score 30.1, $1.15 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.1 Pro Preview: score 29.7, $0.67 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMiniMax-M3: score 29.2, $0.51 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Medium): score 27.6, $1.13 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Low): score 27.5, $0.14 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQuasar 438B (Max, Based on GLM-5.2): score 26.7, $2.02 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTApodex 1.1: score 26.4, $0.46 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Low): score 26.2, $1.05 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.5 Instant (June 2026): score 26.0, $0.69 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTKimi K2.7 Code: score 25.8, $0.54 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTInkling Small: score 25.7, $0.089 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTHy3: score 25.3, $0.072 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.7 Plus: score 25.2, $0.22 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4.1 Flash (Non-reasoning): score 24.7, $0.15 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTSolar Mini 4: score 24.1, $0.36 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNemotron 3 Ultra 550B A55B (Reasoning): score 22.9, $0.60 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGemini 3.5 Flash-Lite: score 22.2, $0.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-5.6 Terra (Non-reasoning): score 20.8, $0.14 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTDeepSeek V4 Pro 0813 (Non-reasoning): score 20.4, $0.48 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.8 27B (Non-reasoning): score 20.2, $2.49 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTLongCat 2.0: score 19.1, $0.059 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Non-reasoning): score 18.5, $0.011 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.5 397B A17B (Reasoning): score 18.4, $0.47 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.6 35B A3B (Reasoning): score 18.2, $0.48 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMuse Glimmer (High): score 17.5, $0.057 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude 4.5 Haiku (Reasoning): score 16.9, $0.28 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTRing-2.6-1T: score 16.6, $0.29 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.5 122B A10B (Reasoning): score 15.6, $0.32 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMistral Medium 3.5: score 14.2, $0.50 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNemotron 3.5 Lightning: score 12.9, $0.093 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNemotron 3 Super 120B A12B (Reasoning): score 12.8, $1.64 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMercury 2.5: score 12.3, $0.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTgpt-oss-120b (High): score 11.6, $0.11 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMistral Small 4 (Reasoning): score 11.3, $0.015 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3.5 9B (Reasoning): score 11.2, $0.21 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGranite 4.2 8B: score 11.1, $0.024 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTTrinity Large Thinking: score 10.8, $0.12 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMistral Large 3: score 9.3, $0.031 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTQwen3 Coder Next: score 9.2, $0.55 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGranite 4.2 3B: score 9.1, $0.0060 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTgpt-oss-20b (High): score 9.0, $0.012 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTNVIDIA Nemotron 3 Nano 30B A3B (Reasoning): score 8.9, $0.017 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTCeleris-1: score 6.3, $0.050 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMinistral 3 14B: score 6.0, $0.019 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMinistral 3 8B: score 5.5, $0.011 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMinistral 3 3B: score 4.8, $0.0078 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback): score 57.6, $5.98 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback): score 56.0, $3.46 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTClaude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback): score 53.6, $1.82 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Max): score 51.8, $0.72 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Xhigh): score 51.0, $0.39 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (High): score 50.2, $0.32 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Medium): score 47.8, $0.21 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMiMo-V2.6-Pro: score 46.3, $0.13 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6.1 Sol (Low): score 42.1, $0.13 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Max): score 38.1, $0.068 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTMiMo-V2.6-Flash: score 37.9, $0.062 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Xhigh): score 34.6, $0.042 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (High): score 32.9, $0.029 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Medium): score 29.9, $0.017 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 Luna (Low): score 21.5, $0.0045 per task · Artificial Analysis, read 4 Oct 2026, 07:09 CESTGPT-6 LunaMiMo-V2.6-ProGPT-6.1 SolOpus 5.5scorecost per task, log scale
Artificial Analysis, LLM Leaderboard, read 4 October 2026. Rows the page marks deprecated, and current rows it gives no cost per task or a cost of zero, are left out. The line and the choice of rows are ours.
The numbers in this chart
scorecost per task
Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)57.6$5.98
Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)56.0$7.67
Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback)56.0$3.46
Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback)53.6$1.82
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)53.4$7.63
Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)53.2$5.98
GPT-6 Astra (Max)52.7$3.26
Gemini 4 Argon (High)52.6$1.99
GPT-6 Astra (Xhigh)52.4$2.31
Claude Sonnet 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback)51.9$2.75
GPT-6.1 Sol (Max)51.8$0.72
Claude Opus 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)51.2$1.34
Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)51.2$3.91
GPT-6.1 Sol (Xhigh)51.0$0.39
GPT-6 Astra (High)50.9$1.73
GPT-6.1 Sol (High)50.2$0.32
GPT-6 Astra (Medium)49.6$1.54
Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback)48.9$2.98
Muse Spark 1.3 (Max)48.1$1.60
GPT-6.1 Sol (Medium)47.8$0.21
Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback)46.8$2.37
Claude Sonnet 5.5 (Adaptive Reasoning, High Effort, Default Fallback)46.8$1.12
Grok 4.7 (Xhigh)46.4$3.74
Grok 4.7 (High)46.3$2.73
MiMo-V2.6-Pro46.3$0.13
GPT-6 Astra (Low)45.8$0.82
Qwen3.8 Max (0902)45.4$5.41
Muse Spark 1.3 (Xhigh)45.1$1.37
GLM-5.3 (Max)44.8$2.01
Step 5 Preview43.7$0.72
Kimi K3 (Max)43.6$2.00
Claude Opus 5.5 (Adaptive Reasoning, Low Effort, Default Fallback)42.3$0.55
Grok 4.7 (Low)42.2$1.25
GPT-6.1 Sol (Low)42.1$0.13
GPT-5.6 Terra (Max)42.1$1.40
GLM 5.3 Flash41.8$0.25
Gemini 3.8 Flash (High)40.9$1.24
Claude Sonnet 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)40.8$0.59
Qwen3.8 2.4T A95B39.9$2.16
Qwen3.8-Flash-Next39.8$0.37
Gemini 3.8 Flash (Medium)39.8$0.93
DeepSeek V4.1 Flash (Max)39.5$0.27
GPT-6 Luna (Max)38.1$0.068
GPT-5.6 Terra (Xhigh)38.0$0.63
MiMo-V2.6-Flash37.9$0.062
DeepSeek V4 Pro 0813 (Max)36.0$0.67
Claude Sonnet 5.5 (Adaptive Reasoning, Low Effort, Default Fallback)35.9$0.42
DeepSeek V4 Flash Vision (Max)34.8$0.31
GPT-6 Luna (Xhigh)34.6$0.042
GLM-5.3 (Low)34.3$0.85
GPT-5.6 Terra (High)34.2$0.34
Qwen3.8 27B (Xhigh)33.7$1.01
GPT-6 Luna (High)32.9$0.029
GPT-5.6 Terra (Medium)30.1$0.18
Kimi K3 (Low)30.1$1.15
GPT-6 Luna (Medium)29.9$0.017
Gemini 3.1 Pro Preview29.7$0.67
MiniMax-M329.2$0.51
Qwen3.8 27B (Medium)27.6$1.13
GPT-5.6 Terra (Low)27.5$0.14
Quasar 438B (Max, Based on GLM-5.2)26.7$2.02
Apodex 1.126.4$0.46
Qwen3.8 27B (Low)26.2$1.05
GPT-5.5 Instant (June 2026)26.0$0.69
Kimi K2.7 Code25.8$0.54
Inkling Small25.7$0.089
Hy325.3$0.072
Qwen3.7 Plus25.2$0.22
DeepSeek V4.1 Flash (Non-reasoning)24.7$0.15
Solar Mini 424.1$0.36
Nemotron 3 Ultra 550B A55B (Reasoning)22.9$0.60
Gemini 3.5 Flash-Lite22.2$0.12
GPT-6 Luna (Low)21.5$0.0045
GPT-5.6 Terra (Non-reasoning)20.8$0.14
DeepSeek V4 Pro 0813 (Non-reasoning)20.4$0.48
Qwen3.8 27B (Non-reasoning)20.2$2.49
LongCat 2.019.1$0.059
GPT-6 Luna (Non-reasoning)18.5$0.011
Qwen3.5 397B A17B (Reasoning)18.4$0.47
Qwen3.6 35B A3B (Reasoning)18.2$0.48
Muse Glimmer (High)17.5$0.057
Claude 4.5 Haiku (Reasoning)16.9$0.28
Ring-2.6-1T16.6$0.29
Qwen3.5 122B A10B (Reasoning)15.6$0.32
Mistral Medium 3.514.2$0.50
Nemotron 3.5 Lightning12.9$0.093
Nemotron 3 Super 120B A12B (Reasoning)12.8$1.64
Mercury 2.512.3$0.12
gpt-oss-120b (High)11.6$0.11
Mistral Small 4 (Reasoning)11.3$0.015
Qwen3.5 9B (Reasoning)11.2$0.21
Granite 4.2 8B11.1$0.024
Trinity Large Thinking10.8$0.12
Mistral Large 39.3$0.031
Qwen3 Coder Next9.2$0.55
Granite 4.2 3B9.1$0.0060
gpt-oss-20b (High)9.0$0.012
NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)8.9$0.017
Celeris-16.3$0.050
Ministral 3 14B6.0$0.019
Ministral 3 8B5.5$0.011
Ministral 3 3B4.8$0.0078

Marked rows are the ones no other row beats on both.

Gemini 4 Argon, new on the list since our first read on 30 September, isn't among them: it scores 52.6 for $1.99 a task at Google's introductory price, and Opus 5.5 at high scores 1.0 points more for 8% less. No setting of Sonnet 5.5, Fable 5.1 or GPT-6 Astra is on the line either.

Every score and cost here is Artificial Analysis's. What we added is the selection and the arithmetic, and both rest on one aggregate score read on one morning.

How the line is drawn

Artificial Analysis publishes many figures for each model, and this article uses two of them: its Intelligence Index, one score across its tests, and a cost per task in US dollars. Its leaderboard lists each reasoning setting as a row of its own, so GPT-6.1 Sol at low and GPT-6.1 Sol at max are two entries.

The page's data holds 688 rows. We took the ones it doesn't mark as deprecated and that have a score and a cost per task above zero, which leaves 102. The page gives a cost of zero for 5 more current rows, and we left those out, since a zero can't be compared with a price. Another 152 current rows have a score but no cost per task on the page, so they can't be placed on the line either. Then we kept every row that scores higher than all the rows costing the same or less. That leaves 15.

Read the chart at the top of the page from left to right. At any cost along the bottom, the line's height is the best score on the list that costs that much or less. The hollow points are the settings you could pick instead, each beaten by something on or under the line.

Where the extra points get expensive

The bottom of the line belongs to GPT-6 Luna, whose five settings from low to max go from $0.0045 to $0.068 a task, with Xiaomi's MiMo-V2.6-Flash between its xhigh and max settings. Then the line moves through GPT-6.1 Sol and MiMo-V2.6-Pro. MiMo-V2.6-Pro costs 2% more per task than GPT-6.1 Sol at low and scores 4.2 points higher. Both MiMo models are open-weights models, the only ones on the line.

GPT-6.1 Sol covers the middle at every setting it has. From low to max its score rises from 42.1 to 51.8, and its cost goes from $0.13 to $0.72. The last step costs the most for the least, since going from xhigh to max adds 0.8 points for 84% more per task.

Above Sol's best score, the only rows on the line are Claude Opus 5.5 at high, xhigh and max. The first of them adds 1.7 points over Sol at max for 2.5× the cost. Opus at max scores 5.8 points more than Sol at max and costs 8.3× as much. Opus's own climb costs a lot too: high to max adds 4.0 points at 3.3× the cost.

The rows on the line, cheapest first

Each step is against the row above it.

settingscorecost per taskpoints morecost multiple
GPT-6 Luna, low21.5$0.0045––
GPT-6 Luna, medium29.9$0.017+8.43.87×
GPT-6 Luna, high32.9$0.029+3.01.66×
GPT-6 Luna, xhigh34.6$0.042+1.61.45×
MiMo-V2.6-Flash37.9$0.062+3.31.47×
GPT-6 Luna, max38.1$0.068+0.21.09×
GPT-6.1 Sol, low42.1$0.13+4.01.93×
MiMo-V2.6-Pro46.3$0.13+4.21.02×
GPT-6.1 Sol, medium47.8$0.21+1.51.60×
GPT-6.1 Sol, high50.2$0.32+2.51.49×
GPT-6.1 Sol, xhigh51.0$0.39+0.81.23×
GPT-6.1 Sol, max51.8$0.72+0.81.84×
Opus 5.5, high53.6$1.82+1.72.52×
Opus 5.5, xhigh56.0$3.46+2.41.90×
Opus 5.5, max57.6$5.98+1.61.73×
Artificial Analysis's scores and costs; the last two columns are our arithmetic, worked out from the page's unrounded figures, so a step can differ by a tenth from the rounded scores shown. Marked: the step from GPT-6.1 Sol's highest setting to Opus 5.5's lowest one on the line.

Settings that cost more for less

Every row off the line has another that costs no more and scores at least as high. The table puts some of the familiar ones next to theirs.

Beaten settings, next to the cheapest row that scores as high

The best current row from several labs that isn't on the line, plus Fable 5.1, Opus 5.5 at medium and Gemini 3.8 Flash at high, each against the cheapest row that scores at least as high.

settingscorecost per taskcheapest row scoring as highits scoreits costits cost as a share
Sonnet 5.5, max56.0$7.67Opus 5.5, max57.6$5.9878%
Fable 5.1, max53.4$7.63Opus 5.5, high53.6$1.8224%
GPT-6 Astra, max52.7$3.26Opus 5.5, high53.6$1.8256%
Gemini 4 Argon, high52.6$1.99Opus 5.5, high53.6$1.8292%
Opus 5.5, medium51.2$1.34GPT-6.1 Sol, max51.8$0.7254%
Muse Spark 1.3, max48.1$1.60GPT-6.1 Sol, high50.2$0.3220%
Grok 4.7, xhigh46.4$3.74GPT-6.1 Sol, medium47.8$0.216%
Qwen3.8 Max45.4$5.41MiMo-V2.6-Pro46.3$0.132%
GLM-5.3, max44.8$2.01MiMo-V2.6-Pro46.3$0.137%
Kimi K3, max43.6$2.00MiMo-V2.6-Pro46.3$0.137%
Gemini 3.8 Flash, high40.9$1.24GPT-6.1 Sol, low42.1$0.1311%
DeepSeek V4.1 Flash, max39.5$0.27GPT-6.1 Sol, low42.1$0.1349%
Artificial Analysis's scores and costs; the matching and the last column are ours.

The Claude rows are the ones a Claude user chooses between. Fable 5.1 at max scores 53.4 for $7.63 a task, and Opus 5.5 at high scores more for 24% of that. Opus 5.5 at medium costs more than GPT-6.1 Sol at max, which scores higher.

Sonnet 5.5 at max and Opus 5.5 at xhigh show the same score to one decimal, 56.0. In the page's unrounded data Sonnet's is a little higher, so the cheapest row that scores at least as high as Sonnet at max is Opus at max, at 78% of Sonnet's cost. Opus at xhigh costs $3.46 a task against Sonnet's $7.67.

What changed since 30 September

We first read the leaderboard for this article on 30 September and didn't publish that version. Between the two reads, 4 rows became current with a score and a price, 6 left the current list, and 9 that are on both lists changed their score by at least 0.05 points or their cost per task by at least 1%. Sonnet 5.5 at max and at xhigh moved by less than that, so we don't count them.

The new rows include Gemini 4 Argon. It scores 52.6 for $1.99 a task, 0.7 points above GPT-6.1 Sol at max for 2.7× its cost. Opus 5.5 at high scores 1.0 points more than Argon and costs 8% less, so Argon arrived off the line. The other new rows are Sonnet 5.5 at low, Grok 4.7 at low and Solar Mini 4, none of them on the line either.

The rows that left are all 6 settings of GPT-6 Sol, which the page now marks as deprecated. None of them was on the line.

GPT-6 Luna's scores rose at every reasoning setting, from +0.5 points at medium to +0.9 points at max, with its costs about where they were. That rise is what put Luna at max on the line, ahead of MiMo-V2.6-Flash, so the line has 15 rows where it had 14. No row left the line.

What this doesn't tell you

It is one aggregate score. A model off this line can still be the right one for a particular job, and the ranking can change test by test. Our GPT-6.1 Sol video goes through Sol and Opus 5.5 test by test, and our Gemini 4 Argon video does the same for Argon.

We wouldn't choose between two rows on a gap of about a point, which is the size of the gap between Argon and Opus at high. What separates them in this comparison is the cost, and for Argon that is an introductory price: Artificial Analysis prices it at Google's $2 and $10 per million input and output tokens, and Google says $4 and $20 will apply after the introductory period. With every part of the price doubled, Argon would cost 2.2× as much per task as Opus at high. Google doesn't say what cached input will cost then, and our Gemini 4 Argon video works through that.

The cost is Artificial Analysis's cost per task on its own tests, in dollars. On a Claude or ChatGPT subscription you pay in quota instead, and this list doesn't measure that.

An older GPT-5.6 Luna row would sit on the line if we had counted rows the page marks as deprecated. We left it out because the question is what to pick today.

Artificial Analysis updates its figures, and Luna's rise between our two reads shows how a few days can move a row on or off the line. Every figure here is dated in the numbers table below.

Every number

These are all 498 figures behind this article, grouped by whose they are, with the page each came from and when we read it. Figures marked ⟳ can move. When a re-read finds a change, the new value shows next to the one we first published.

Artificial Analysis: score and cost per task, every current model and setting with a price

scorecost per task
Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)57.6$5.98
Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)56.0$7.67
Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback)56.0$3.46
Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback)53.6$1.82
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)53.4$7.63
Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)53.2$5.98
GPT-6 Astra (Max)52.7$3.26
Gemini 4 Argon (High)52.6$1.99
GPT-6 Astra (Xhigh)52.4$2.31
Claude Sonnet 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback)51.9$2.75
GPT-6.1 Sol (Max)51.8$0.72
Claude Opus 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)51.2$1.34
Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)51.2$3.91
GPT-6.1 Sol (Xhigh)51.0$0.39
GPT-6 Astra (High)50.9$1.73
GPT-6.1 Sol (High)50.2$0.32
GPT-6 Astra (Medium)49.6$1.54
Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback)48.9$2.98
Muse Spark 1.3 (Max)48.1$1.60
GPT-6.1 Sol (Medium)47.8$0.21
Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback)46.8$2.37
Claude Sonnet 5.5 (Adaptive Reasoning, High Effort, Default Fallback)46.8$1.12
Grok 4.7 (Xhigh)46.4$3.74
Grok 4.7 (High)46.3$2.73
MiMo-V2.6-Pro46.3$0.13
GPT-6 Astra (Low)45.8$0.82
Qwen3.8 Max (0902)45.4$5.41
Muse Spark 1.3 (Xhigh)45.1$1.37
GLM-5.3 (Max)44.8$2.01
Step 5 Preview43.7$0.72
Kimi K3 (Max)43.6$2.00
Claude Opus 5.5 (Adaptive Reasoning, Low Effort, Default Fallback)42.3$0.55
Grok 4.7 (Low)42.2$1.25
GPT-6.1 Sol (Low)42.1$0.13
GPT-5.6 Terra (Max)42.1$1.40
GLM 5.3 Flash41.8$0.25
Gemini 3.8 Flash (High)40.9$1.24
Claude Sonnet 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)40.8$0.59
Qwen3.8 2.4T A95B39.9$2.16
Qwen3.8-Flash-Next39.8$0.37
Gemini 3.8 Flash (Medium)39.8$0.93
DeepSeek V4.1 Flash (Max)39.5$0.27
GPT-6 Luna (Max)38.1$0.068
GPT-5.6 Terra (Xhigh)38.0$0.63
MiMo-V2.6-Flash37.9$0.062
DeepSeek V4 Pro 0813 (Max)36.0$0.67
Claude Sonnet 5.5 (Adaptive Reasoning, Low Effort, Default Fallback)35.9$0.42
DeepSeek V4 Flash Vision (Max)34.8$0.31
GPT-6 Luna (Xhigh)34.6$0.042
GLM-5.3 (Low)34.3$0.85
GPT-5.6 Terra (High)34.2$0.34
Qwen3.8 27B (Xhigh)33.7$1.01
GPT-6 Luna (High)32.9$0.029
GPT-5.6 Terra (Medium)30.1$0.18
Kimi K3 (Low)30.1$1.15
GPT-6 Luna (Medium)29.9$0.017
Gemini 3.1 Pro Preview29.7$0.67
MiniMax-M329.2$0.51
Qwen3.8 27B (Medium)27.6$1.13
GPT-5.6 Terra (Low)27.5$0.14
Quasar 438B (Max, Based on GLM-5.2)26.7$2.02
Apodex 1.126.4$0.46
Qwen3.8 27B (Low)26.2$1.05
GPT-5.5 Instant (June 2026)26.0$0.69
Kimi K2.7 Code25.8$0.54
Inkling Small25.7$0.089
Hy325.3$0.072
Qwen3.7 Plus25.2$0.22
DeepSeek V4.1 Flash (Non-reasoning)24.7$0.15
Solar Mini 424.1$0.36
Nemotron 3 Ultra 550B A55B (Reasoning)22.9$0.60
Gemini 3.5 Flash-Lite22.2$0.12
GPT-6 Luna (Low)21.5$0.0045
GPT-5.6 Terra (Non-reasoning)20.8$0.14
DeepSeek V4 Pro 0813 (Non-reasoning)20.4$0.48
Qwen3.8 27B (Non-reasoning)20.2$2.49
LongCat 2.019.1$0.059
GPT-6 Luna (Non-reasoning)18.5$0.011
Qwen3.5 397B A17B (Reasoning)18.4$0.47
Qwen3.6 35B A3B (Reasoning)18.2$0.48
Muse Glimmer (High)17.5$0.057
Claude 4.5 Haiku (Reasoning)16.9$0.28
Ring-2.6-1T16.6$0.29
Qwen3.5 122B A10B (Reasoning)15.6$0.32
Mistral Medium 3.514.2$0.50
Nemotron 3.5 Lightning12.9$0.093
Nemotron 3 Super 120B A12B (Reasoning)12.8$1.64
Mercury 2.512.3$0.12
gpt-oss-120b (High)11.6$0.11
Mistral Small 4 (Reasoning)11.3$0.015
Qwen3.5 9B (Reasoning)11.2$0.21
Granite 4.2 8B11.1$0.024
Trinity Large Thinking10.8$0.12
Mistral Large 39.3$0.031
Qwen3 Coder Next9.2$0.55
Granite 4.2 3B9.1$0.0060
gpt-oss-20b (High)9.0$0.012
NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)8.9$0.017
Celeris-16.3$0.050
Ministral 3 14B6.0$0.019
Ministral 3 8B5.5$0.011
Ministral 3 3B4.8$0.0078

Artificial Analysis, our 30 September read: score and cost per task, every current model and setting with a price

Artificial Analysis · Artificial Analysis, LLM Leaderboard, our earlier read (the page's own data) · read 30 Sep 2026, 21:51 CEST
scorecost per task
Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)57.6$5.98
Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback)56.0$3.46
Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)56.0$7.60
Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback)53.6$1.82
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)53.4$7.63
Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)53.2$5.98
GPT-6 Astra (max)52.7$3.26
GPT-6 Astra (xhigh)52.4$2.31
Claude Sonnet 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback)51.9$2.74
GPT-6.1 Sol (max)51.8$0.72
Claude Opus 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)51.2$1.34
Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)51.2$3.91
GPT-6.1 Sol (xhigh)51.0$0.39
GPT-6 Astra (high)50.9$1.73
GPT-6.1 Sol (high)50.2$0.32
GPT-6 Astra (medium)49.6$1.54
Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback)48.9$2.98
Muse Spark 1.3 (max)48.1$1.60
GPT-6.1 Sol (medium)47.8$0.21
GPT-6 Sol (max)47.5$1.05
Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback)46.8$2.37
Claude Sonnet 5.5 (Adaptive Reasoning, High Effort, Default Fallback)46.7$1.08
Grok 4.7 (xhigh)46.4$3.74
Grok 4.7 (high)46.3$2.73
MiMo-V2.6-Pro46.3$0.13
GPT-6 Astra (low)45.8$0.82
Qwen3.8 Max (0902)45.4$5.41
Muse Spark 1.3 (xhigh)45.1$1.37
GLM-5.3 (max)44.8$2.01
GPT-6 Sol (xhigh)44.1$0.52
Step 5 Preview43.7$0.72
Kimi K3 (max)43.6$2.00
GPT-6 Sol (high)42.8$0.38
Claude Opus 5.5 (Adaptive Reasoning, Low Effort, Default Fallback)42.3$0.55
GPT-6.1 Sol (low)42.1$0.13
GPT-5.6 Terra (max)42.1$1.40
GLM 5.3 Flash41.8$0.25
Gemini 3.8 Flash (high)40.9$1.24
Claude Sonnet 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)40.7$0.59
Qwen3.8 2.4T A95B39.9$2.16
Qwen3.8-Flash-Next39.8$0.37
GPT-6 Sol (medium)39.8$0.25
Gemini 3.8 Flash (medium)39.8$0.93
DeepSeek V4.1 Flash (Reasoning, Max Effort)39.5$0.27
GPT-5.6 Terra (xhigh)38.0$0.63
MiMo-V2.6-Flash37.9$0.062
GPT-6 Luna (max)37.3$0.068
DeepSeek V4 Pro 0813 (Reasoning, Max Effort)36.0$0.67
DeepSeek V4 Flash Vision (Reasoning, Max Effort)34.8$0.31
GLM-5.3 (low)34.3$0.85
GPT-5.6 Terra (high)34.2$0.34
GPT-6 Sol (low)33.9$0.13
GPT-6 Luna (xhigh)33.9$0.042
Qwen3.8 27B (xhigh)33.7$1.01
GPT-6 Luna (high)32.1$0.029
GPT-5.6 Terra (medium)30.1$0.18
Kimi K3 (low)30.1$1.15
Gemini 3.1 Pro Preview29.7$0.67
GPT-6 Luna (medium)29.5$0.018
MiniMax-M329.2$0.51
GPT-6 Sol (Non-reasoning)28.1$0.33
Qwen3.8 27B (medium)27.6$1.13
GPT-5.6 Terra (low)27.5$0.14
Quasar 438B (max, based on GLM-5.2)26.7$2.02
Apodex 1.126.4$0.46
Qwen3.8 27B (low)26.2$1.05
GPT-5.5 Instant (June 2026)26.0$0.69
Kimi K2.7 Code25.8$0.54
Inkling Small25.7$0.089
Hy325.3$0.072
Qwen3.7 Plus25.2$0.22
DeepSeek V4.1 Flash (Non-Reasoning)24.7$0.15
Nemotron 3 Ultra 550B A55B (Reasoning)22.9$0.60
Gemini 3.5 Flash-Lite22.2$0.12
GPT-6 Luna (low)20.9$0.0045
GPT-5.6 Terra (Non-reasoning)20.8$0.14
DeepSeek V4 Pro 0813 (Non-reasoning)20.4$0.48
Qwen3.8 27B (Non-reasoning)20.2$2.49
LongCat 2.019.1$0.059
Qwen3.5 397B A17B (Reasoning)18.4$0.47
GPT-6 Luna (Non-reasoning)18.3$0.011
Qwen3.6 35B A3B (Reasoning)18.2$0.48
Muse Glimmer (high)17.5$0.057
Claude 4.5 Haiku (Reasoning)16.9$0.28
Ring-2.6-1T16.6$0.29
Qwen3.5 122B A10B (Reasoning)15.6$0.32
Mistral Medium 3.514.2$0.50
Nemotron 3.5 Lightning12.9$0.10
Nemotron 3 Super 120B A12B (Reasoning)12.8$1.64
Mercury 2.512.3$0.12
gpt-oss-120b (high)11.6$0.11
Mistral Small 4 (Reasoning)11.3$0.015
Qwen3.5 9B (Reasoning)11.2$0.21
Granite 4.2 8B11.1$0.024
Trinity Large Thinking10.8$0.12
Mistral Large 39.3$0.031
Qwen3 Coder Next9.2$0.55
Granite 4.2 3B9.1$0.0060
gpt-oss-20b (high)9.0$0.012
NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)8.9$0.017
Celeris-16.3$0.050
Ministral 3 14B6.0$0.019
Ministral 3 8B5.5$0.011
Ministral 3 3B4.8$0.0078

Artificial Analysis

Rows on the leaderboard not marked deprecated that have a score and a cost per task above zero (our count)102 ⟳
Rows in the page's data, every model and setting, deprecated ones included (our count)688 ⟳
Rows not marked deprecated that the page gives a cost per task of zero, left out (our count)5 ⟳
Rows not marked deprecated that have a score but no cost per task on the page, left out (our count)152 ⟳

Model Fatigue, from Artificial Analysis's figures

Our arithmetic on Artificial Analysis's figures (derived/frontier.py) · from the read of 4 Oct 2026, 07:09 CEST
Rows no other current row beats on both score and cost per task (our arithmetic)15 ⟳
Score gained moving up the line, GPT-6 Luna (Low) to GPT-6 Luna (Medium) (our arithmetic)+8.4
Cost per task multiple for that step, GPT-6 Luna (Low) to GPT-6 Luna (Medium) (our arithmetic)3.87×
Score gained moving up the line, GPT-6 Luna (Medium) to GPT-6 Luna (High) (our arithmetic)+3.0
Cost per task multiple for that step, GPT-6 Luna (Medium) to GPT-6 Luna (High) (our arithmetic)1.66×
Score gained moving up the line, GPT-6 Luna (High) to GPT-6 Luna (Xhigh) (our arithmetic)+1.6
Cost per task multiple for that step, GPT-6 Luna (High) to GPT-6 Luna (Xhigh) (our arithmetic)1.45×
Score gained moving up the line, GPT-6 Luna (Xhigh) to MiMo-V2.6-Flash (our arithmetic)+3.3
Cost per task multiple for that step, GPT-6 Luna (Xhigh) to MiMo-V2.6-Flash (our arithmetic)1.47×
Score gained moving up the line, MiMo-V2.6-Flash to GPT-6 Luna (Max) (our arithmetic)+0.2
Cost per task multiple for that step, MiMo-V2.6-Flash to GPT-6 Luna (Max) (our arithmetic)1.09×
Score gained moving up the line, GPT-6 Luna (Max) to GPT-6.1 Sol (Low) (our arithmetic)+4.0
Cost per task multiple for that step, GPT-6 Luna (Max) to GPT-6.1 Sol (Low) (our arithmetic)1.93×
Score gained moving up the line, GPT-6.1 Sol (Low) to MiMo-V2.6-Pro (our arithmetic)+4.2
Cost per task multiple for that step, GPT-6.1 Sol (Low) to MiMo-V2.6-Pro (our arithmetic)1.02×
Score gained moving up the line, MiMo-V2.6-Pro to GPT-6.1 Sol (Medium) (our arithmetic)+1.5
Cost per task multiple for that step, MiMo-V2.6-Pro to GPT-6.1 Sol (Medium) (our arithmetic)1.60×
Score gained moving up the line, GPT-6.1 Sol (Medium) to GPT-6.1 Sol (High) (our arithmetic)+2.5
Cost per task multiple for that step, GPT-6.1 Sol (Medium) to GPT-6.1 Sol (High) (our arithmetic)1.49×
Score gained moving up the line, GPT-6.1 Sol (High) to GPT-6.1 Sol (Xhigh) (our arithmetic)+0.8
Cost per task multiple for that step, GPT-6.1 Sol (High) to GPT-6.1 Sol (Xhigh) (our arithmetic)1.23×
Score gained moving up the line, GPT-6.1 Sol (Xhigh) to GPT-6.1 Sol (Max) (our arithmetic)+0.8
Cost per task multiple for that step, GPT-6.1 Sol (Xhigh) to GPT-6.1 Sol (Max) (our arithmetic)1.84×
Score gained moving up the line, GPT-6.1 Sol (Max) to Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) (our arithmetic)+1.7
Cost per task multiple for that step, GPT-6.1 Sol (Max) to Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) (our arithmetic)2.52×
Score gained moving up the line, Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) to Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback) (our arithmetic)+2.4
Cost per task multiple for that step, Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) to Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback) (our arithmetic)1.90×
Score gained moving up the line, Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback) to Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) (our arithmetic)+1.6
Cost per task multiple for that step, Claude Opus 5.5 (Adaptive Reasoning, Xhigh Effort, Default Fallback) to Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) (our arithmetic)1.73×
Cost per task of Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) as a share of Claude Sonnet 5.5 (Adaptive Reasoning, Max Effort, Default Fallback)'s (our arithmetic)78%
Cost per task of Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) as a share of Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)'s (our arithmetic)24%
Cost per task of Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) as a share of GPT-6 Astra (Max)'s (our arithmetic)56%
Cost per task of Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) as a share of Gemini 4 Argon (High)'s (our arithmetic)92%
Cost per task of GPT-6.1 Sol (Max) as a share of Claude Opus 5.5 (Adaptive Reasoning, Medium Effort, Default Fallback)'s (our arithmetic)54%
Cost per task of GPT-6.1 Sol (High) as a share of Muse Spark 1.3 (Max)'s (our arithmetic)20%
Cost per task of GPT-6.1 Sol (Medium) as a share of Grok 4.7 (Xhigh)'s (our arithmetic)6%
Cost per task of MiMo-V2.6-Pro as a share of Qwen3.8 Max (0902)'s (our arithmetic)2%
Cost per task of MiMo-V2.6-Pro as a share of GLM-5.3 (Max)'s (our arithmetic)7%
Cost per task of MiMo-V2.6-Pro as a share of Kimi K3 (Max)'s (our arithmetic)7%
Cost per task of GPT-6.1 Sol (Low) as a share of Gemini 3.8 Flash (High)'s (our arithmetic)11%
Cost per task of GPT-6.1 Sol (Low) as a share of DeepSeek V4.1 Flash (Max)'s (our arithmetic)49%
Score difference, Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) over GPT-6.1 Sol (Max) (our arithmetic)5.8 points
Cost per task multiple, Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) over GPT-6.1 Sol (Max) (our arithmetic)8.3×
Cost per task change, Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) against GPT-6.1 Sol (Max) (our arithmetic)+726%
Score difference, Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) over GPT-6.1 Sol (Max) (our arithmetic)1.7 points
Cost per task multiple, Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) over GPT-6.1 Sol (Max) (our arithmetic)2.5×
Cost per task change, Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) against GPT-6.1 Sol (Max) (our arithmetic)+152%
Score difference, GPT-6.1 Sol (Max) over GPT-6.1 Sol (Xhigh) (our arithmetic)0.8 points
Cost per task multiple, GPT-6.1 Sol (Max) over GPT-6.1 Sol (Xhigh) (our arithmetic)1.8×
Cost per task change, GPT-6.1 Sol (Max) against GPT-6.1 Sol (Xhigh) (our arithmetic)+84%
Score difference, MiMo-V2.6-Pro over GPT-6.1 Sol (Low) (our arithmetic)4.2 points
Cost per task multiple, MiMo-V2.6-Pro over GPT-6.1 Sol (Low) (our arithmetic)1.0×
Cost per task change, MiMo-V2.6-Pro against GPT-6.1 Sol (Low) (our arithmetic)+2%
Score difference, Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) over Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) (our arithmetic)4.0 points
Cost per task multiple, Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) over Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) (our arithmetic)3.3×
Cost per task change, Claude Opus 5.5 (Adaptive Reasoning, Max Effort, Default Fallback) against Claude Opus 5.5 (Adaptive Reasoning, High Effort, Default Fallback) (our arithmetic)+228%
Score difference, GPT-6 Luna (Xhigh) over GPT-6 Luna (Low) (our arithmetic)13.0 points
Cost per task multiple, GPT-6 Luna (Xhigh) over GPT-6 Luna (Low) (our arithmetic)9.3×
Cost per task change, GPT-6 Luna (Xhigh) against GPT-6 Luna (Low) (our arithmetic)+834%
Score change between the two reads, GPT-6 Luna (Low) (our arithmetic)+0.6 points
Score change between the two reads, GPT-6 Luna (Medium) (our arithmetic)+0.5 points
Score change between the two reads, GPT-6 Luna (High) (our arithmetic)+0.8 points
Score change between the two reads, GPT-6 Luna (Xhigh) (our arithmetic)+0.7 points
Score change between the two reads, GPT-6 Luna (Max) (our arithmetic)+0.9 points
Score difference, Claude Opus 5.5 at high over Gemini 4 Argon (our arithmetic)1.0 points
How much less Claude Opus 5.5 at high costs per task than Gemini 4 Argon (our arithmetic)8%
Gemini 4 Argon's cost per task with every part of its price doubled, as a multiple of Claude Opus 5.5 at high's (our arithmetic)2.2×
Score difference, Gemini 4 Argon over GPT-6.1 Sol at max (our arithmetic)0.7 points
Cost per task multiple, Gemini 4 Argon over GPT-6.1 Sol at max (our arithmetic)2.7×

Model Fatigue, from Artificial Analysis's figures

Our comparison of the two reads (derived/changes.py) · from the read of 4 Oct 2026, 07:09 CEST
Current priced rows in the 30 September read (our count)104
Rows current and priced on 4 October that weren't on 30 September (our count)4
Rows current and priced on 30 September that aren't on 4 October (our count)6
Of those, rows the 4 October read marks deprecated (our count)6
Rows in both reads whose score moved by 0.05 points or more or whose cost per task moved by 1% or more (our count)9
Smallest score change we counted as a change between the reads (points; our threshold)0.05
Smallest cost per task change we counted as a change between the reads (our threshold)1%
Rows in both reads that changed by less than those thresholds (our count)2
Rows on the line in the 30 September read (our arithmetic)14

Google

Gemini 4 Argon, introductory API price per million input tokens$2
Gemini 4 Argon, introductory API price per million output tokens$10
Gemini 4 Argon, API price per million input tokens after the introductory period$4
Gemini 4 Argon, API price per million output tokens after the introductory period$20

Sources

These are the pages this article draws on. We keep a copy of each page as we read it, so a figure can be checked against what the page said at the time.