Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Loading page content
Back to models
GPT-4o mini
Closed
Share
Add to comparison
Tool use
Overview
Pricing
Evaluations
Evaluation overview
Overall rank
#158
51.5 points · 180 models
Agents rank
#104
56.3 points · 125 models
Data coverage
70 items
6 sources · updated through 2026-07-16
Domain performance
Coding
#134 / 160
60.9
Research
#151 / 174
61.4
Agents
#104 / 125
56.3
Vision
#63 / 65
63.3
Long context
#78 / 82
53.6
Evaluation highlights
All evaluations
BFCL v3 Full
Agents · Qwen3 Technical Report
64%
23
/ 49
MMLU
Reasoning · OpenAI GPT-4.1 Release
82%
25
/ 46
LiveCodeBench v5
Coding · Qwen3 Technical Report
27.9%
23
/ 36
Arena Vision Rating
Multimodal · Arena.ai Vision
1,097.6Elo
91
/ 121
OpenAI MRCR 2-needle 128k
Long context · OpenAI GPT-4.1 Release
24.5%
11
/ 13
IFEval
Instruction following · OpenAI GPT-4.1 Release
78.4%
28
/ 30
Basic information
Context (input / output)
128K / 16K
Release date
2024-07
Input modalities
Text · Image
Output modalities
Text
Availability
OpenAI
Full pricing
Input
$0.15 /1M
Output
$0.6 /1M
Related links
Official site