Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Benchmark details loading
Back to benchmarks
PHYBench
Reasoning
Share
Results for 11 models collected from 1 public source.
Ranking
Sources
#
Model
Score
1
Gemini 3.1 Pro
Google
82.0
%
2
Claude Opus 4.8
Anthropic
79.6
%
3
GPT-5.5
OpenAI
77.5
%
4
Seed2.1 Pro
ByteDance
77.5
%
5
Hunyuan 3
Tencent
77.4
%
6
DeepSeek-V4-Flash
DeepSeek
76.6
%
7
Qwen3.7-Max
Alibaba
76.5
%
8
DeepSeek-V4-Pro
DeepSeek
75.4
%
9
GLM-5.2
Zhipu AI
71.5
%
10
Hunyuan 3 Preview
Tencent
71.4
%
11 models total
1
2