Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Loading page content
Back to models
GPT-4o
Closed
Share
Add to comparison
Tool use
Overview
Pricing
Evaluations
Evaluation overview
Overall rank
#134
59.1 points · 180 models
Agents rank
#76
65 points · 125 models
Data coverage
193 items
10 sources · updated through 2026-07-16
Domain performance
Coding
#114 / 160
64.1
Research
#130 / 174
69
Agents
#76 / 125
65
Vision
#53 / 65
69.9
Long context
#71 / 82
60.4
Evaluation highlights
All evaluations
BFCL v3 Full
Agents · Qwen3 Technical Report
72.5%
4
/ 49
Arena Text Rating
Reasoning · Arena.ai Text · ChatGPT 2025-03-26 mode
1,443.19Elo
53
/ 354
Arena Vision Rating
Multimodal · Arena.ai Vision · ChatGPT 2025-03-26 mode
1,240.83Elo
30
/ 121
Aider Edit
Coding · DeepSeek-V3 Technical Report
72.9%
3
/ 7
LongBench v2
Long context · DeepSeek-V3 Technical Report
48.1%
15
/ 22
IFEval
Instruction following · OpenAI GPT-4.1 Release
81%
26
/ 30
Basic information
Context (input / output)
128K / 16K
Release date
2024-05
Input modalities
Text · Image
Output modalities
Text
Availability
OpenAI
Full pricing
Input
$2.5 /1M
Output
$10.42 /1M
Related links
Official site
Version relationships
Provider variants
ChatGPT 2025-03-26 · Search