Skip to main content
leaderboard
.cn
Toggle Sidebar
Search models or organizations…
⌘K
Explore
Frontier Models
Historical Trends
Local Models
Compare Models
Ranking Studio
Library
Models
Pricing
Organizations
Benchmarks
Data sources
Settings
Toggle Sidebar
leaderboard
.cn
Loading page content
Feedback
Loading page content
Back to models
gpt-oss-120b
Open
Share
Add to comparison
Tool use · Reasoning
Overview
Deployment
Evaluations
Evaluation overview
Overall rank
#123
66.1 points · 202 models
Open-weight rank
#52
69 points · 120 models
Data coverage
15 items
3 sources · updated through 2026-09-02
Domain performance
Coding
#78 / 182
69.6
Research
#120 / 196
74.9
Long context
#73 / 95
69.2
Evaluation highlights
All evaluations
MMLU-Redux
Reasoning · Qwen3.5-9B Technical Report
91%
31
/ 102
LongBench v2
Long context · Qwen3.5-122B-A10B Technical Report
48.2%
19
/ 27
IFEval strict prompt
Instruction following · Qwen3.5-9B Technical Report
88.9%
16
/ 55
C-Eval
Chinese · Qwen3.5-9B Technical Report
76.2%
42
/ 64
Arena Text Rating
Reasoning · Arena.ai Text
1,352.62Elo
177
/ 375
SuperGPQA
Reasoning · Qwen3.5-9B Technical Report
54.6%
31
/ 65
Basic information
Context (input / output)
131K / 131K
Parameters
117B / 5.1B active
Release date
2025-08
License
Apache 2.0
Input modalities
Text
Output modalities
Text
Architecture
MoE
Reliable knowledge cutoff
2024-06
Availability
vLLM · MXFP4
Deployment options
MXFP4
Related links
Official site
HuggingFace