- Context (input / output)
- 1.05M / 131K
- Parameters
- 320B / 18B active
- Release date
- 2026-08
- License
- MIT
- Input modalities
- Text · Image · Video · file
- Output modalities
- Text
- Layers
- 45
- Architecture
- MoE, sparse + linear attention, mHC, IndexPool
- Total parameters
- 320B
- Exact release date
- 2026-08-26
- Availability
- GLM Coding Plan and API
- API model ID
- glm-5.3-flash
- Reasoning mode
- Thinking only (supports low / high / max)
- Weight status
- Open weights released (MIT)
- Evaluation source
- Official 14-row comparison table with 70 scores; base-model table excluded
- Pretraining data
- 30T-token multimodal corpus
- Default reasoning effort
- max (supports low / high / max)