// leaderboard
AI Model Leaderboard
Ranked by benchmark score latency, cost-efficiency, accuracy, coding, and reasoning.
ExploreAI Tools
Best for All
Claude Sonnet 4 โ Anthropic
#Model
๐ฅ95%94%1.1s$3.00
Claude Sonnet 4Anthropic
92
๐ฅ94%92%980ms$2.00
GPT-4.1OpenAI
91
๐ฅ93%93%1.2s$3.00
Claude 3.7 SonnetAnthropic
90
491%89%820ms$2.50
GPT-4oOpenAI
89
592%91%1.3s$3.50
Gemini 2.5 ProGoogle
89
688%91%760ms$0.27
DeepSeek V3DeepSeek
89
788%90%1.1s$1.10
o4-miniOpenAI
88
885%84%550ms$0.35
Gemini 2.5 FlashGoogle
87
987%88%2.1sFree
DeepSeek R1DeepSeekfree
86
1086%85%900ms$2.00
Mistral LargeMistral
86
1182%90%1.2sFree
Qwen3 Coder 480BAlibabafree
85
1282%81%380ms$0.15
GPT-4o miniOpenAI
85
1383%82%450ms$0.25
Claude 3.5 HaikuAnthropic
85
1484%83%1.5sFree
GPT OSS 120BOpenAIfree
84
1580%79%870msFree
Llama 3.3 70BMetafree
82
1680%81%1.1sFree
Nemotron Super 120BNVIDIAfree
82
1778%76%900msFree
Gemma 3 27BGooglefree
81
1877%75%880msFree
MiniMax M2.5MiniMaxfree
80
1997%96%4.2s$15.00
o3OpenAI
80
2075%74%680msFree
GPT OSS 20BOpenAIfree
79
2173%72%520msFree
GLM 4.5 AirZ.aifree
ExploreAI Tools
78
2272%70%500msFree
Gemma 3 12BGooglefree
77
2371%72%450msFree
Nemotron Nano 30BNVIDIAfree
77
2465%63%290msFree
Llama 3.2 3BMetafree
72
Discover more
Scores are composite benchmarks. Figures are representative and updated periodically.
