Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Loading page content
Back to models
gpt-oss-120b
Open
Share
Add to comparison
Tool use · Reasoning
Overview
Deployment
Evaluations
Evaluation overview
Overall rank
#102
65.8 points · 180 models
Open-weight rank
#41
72.3 points · 109 models
Data coverage
15 items
3 sources · updated through 2026-07-16
Domain performance
Coding
#58 / 160
73.2
Research
#97 / 174
76
Long context
#60 / 82
67
Evaluation highlights
All evaluations
MMLU-Redux
Reasoning · Qwen3.5-9B Technical Report
91%
31
/ 102
LongBench v2
Long context · Qwen3.5-122B-A10B Technical Report
48.2%
14
/ 22
IFEval strict prompt
Instruction following · Qwen3.5-9B Technical Report
88.9%
16
/ 55
C-Eval
Chinese · Qwen3.5-9B Technical Report
76.2%
42
/ 64
Arena Text Rating
Reasoning · Arena.ai Text
1,352.62Elo
159
/ 354
MMLU-Pro
Reasoning · Qwen3.5-9B Technical Report
80.8%
45
/ 95
Basic information
Context (input / output)
131K / 131K
Parameters
117B / 5.1B active
Release date
2025-08
License
Apache 2.0
Input modalities
Text
Output modalities
Text
Availability
vLLM · MXFP4
Deployment options
MXFP4
Related links
Official site
HuggingFace