Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Loading page content
Back to models
GPT-4o
Closed
Share
Add to comparison
Tool use
Overview
Pricing
Evaluations
Evaluation overview
Overall rank
#156
59.5 points · 202 models
Agents rank
#96
63.5 points · 148 models
Data coverage
193 items
10 sources · updated through 2026-09-02
Domain performance
Coding
#137 / 182
60.3
Research
#150 / 196
68.5
Agents
#96 / 148
63.5
Vision
#70 / 82
65
Long context
#84 / 95
61.4
Evaluation highlights
All evaluations
BFCL v3 Full
Agents · Qwen3 Technical Report
72.5%
4
/ 49
Arena Text Rating
Reasoning · Arena.ai Text · ChatGPT 2025-03-26 mode
1,443.19Elo
66
/ 375
Arena Vision Rating
Multimodal · Arena.ai Vision · ChatGPT 2025-03-26 mode
1,240.83Elo
41
/ 133
Aider Edit
Coding · DeepSeek-V3 Technical Report
72.9%
3
/ 7
LongBench v2
Long context · DeepSeek-V3 Technical Report
48.1%
20
/ 27
IFEval
Instruction following · OpenAI GPT-4.1 Release
81%
26
/ 30
Basic information
Context (input / output)
128K / 16K
Release date
2024-05
Input modalities
Text · Image
Output modalities
Text
Availability
OpenAI
Full pricing
Input
$2.5 /1M
Output
$10.42 /1M
Related links
Official site
Version relationships
Provider variants
ChatGPT 2025-03-26 · Search