Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Benchmark details loading
Back to benchmarks
AIME 2024 Cons@64
Reasoning
Share
Results for 10 models collected from 1 public source.
Ranking
Sources
#
Model
Score
1
DeepSeek-R1-Distill-Llama-70B
DeepSeek
86.7
%
2
DeepSeek-R1-Distill-Qwen-7B
DeepSeek
83.3
%
3
DeepSeek-R1-Distill-Qwen-32B
DeepSeek
83.3
%
4
DeepSeek-R1-Distill-Qwen-14B
DeepSeek
80.0
%
5
DeepSeek-R1-Distill-Llama-8B
DeepSeek
80.0
%
6
OpenAI o1-mini
OpenAI
80.0
%
7
QwQ-32B-Preview
Alibaba
60.0
%
8
DeepSeek-R1-Distill-Qwen-1.5B
DeepSeek
52.7
%
9
Claude 3.5 Sonnet 20241022
Anthropic
26.7
%
10
GPT-4o
OpenAI
13.4
%
10 models total