Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Loading page content
Back to models
o3
Closed
Share
Add to comparison
Tool use · Reasoning
Overview
Pricing
Evaluations
Evaluation overview
Overall rank
#62
71.6 points · 180 models
Coding rank
#51
74.7 points · 160 models
Data coverage
30 items
3 sources · updated through 2026-07-01
Domain performance
Coding
#51 / 160
74.7
Research
#71 / 174
80.2
Agents
#46 / 125
72.3
Vision
#37 / 65
79.1
Long context
#49 / 82
72.3
Evaluation highlights
All evaluations
Aider Polyglot
Coding · OpenAI GPT-5 Developer Release
79.6%
2
/ 29
TAU2-Airline
Agents · OpenAI GPT-5 Developer Release
64.8%
2
/ 17
AIME 2025
Reasoning · OpenAI GPT-5 Developer Release
88.9%
13
/ 60
Graphwalks BFS 0K-128K
Long context · OpenAI GPT-5 Developer Release
77.3%
4
/ 18
VideoMME long (with subtitles)
Multimodal · OpenAI GPT-5 Developer Release
84.9%
2
/ 8
COLLIE
Instruction following · OpenAI GPT-5 Developer Release
98.4%
4
/ 13
Basic information
Context (input / output)
200K / 100K
Release date
2025-04
Input modalities
Text · Image
Output modalities
Text
Availability
OpenAI
Full pricing
Input
$1 /1M
Output
$4 /1M
Related links
Official site