Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Benchmark details loading
Back to benchmarks
Toolathlon-Verified
Agents
Share
Official website
Results for 24 models collected from 12 public sources.
Ranking
Sources
#
Model
Score
1
GLM-5.3-Flash
Zhipu AI
78.4
%
2
Claude Fable 5
Anthropic
77.9
%
3
Claude Opus 5
Anthropic
77.6
%
4
Kimi K3
Moonshot AI
76.5
%
5
Claude Opus 4.8
Anthropic
76.2
%
6
DeepSeek-V4-Flash-Vision-Exp
DeepSeek
75.9
%
7
Muse Spark 1.1
Meta
75.6
%
8
GPT-5.6 Sol
OpenAI
74.9
%
9
GPT-5.6 Terra
OpenAI
74.9
%
10
DeepSeek-V4-Pro-0813
DeepSeek
74.1
%
24 models total
1
2
3