GLM-5.3 Flash
The GLM-5.3 Flash checkpoint is published by Z.ai for language and reasoning. Its weights are available under MIT; that alone does not establish that its code and all components are open source.
Consider this MoE checkpoint when comparing language and reasoning models with published total and active parameter counts.
320B total and 18B active are different quantities. Runtime support and any image components still need verification.
What is documented
- Checkpoint / version
- zai-org/GLM-5.3-Flash ↗
- Weight license
- MIT ↗
- Task and modality
- language and reasoning · text + image
- Count and basis
- 320B · manufacturer specification · 18B active
- Context
- Not verified for this checkpoint
- Quality evidence
- 71.6 · LiveBench 2026-06-25 overall · source ↗
- Runtime and precision
- No local execution verified. Check the publisher's supported runtime and checkpoint precision; batch and context memory are not measured here.
Run locally: known and unknown
~248 GB generic Q4 planning memory. Formula: parameters × 0.625 bytes, plus 20% working allowance and 8 GB reserve. This is neither measured VRAM nor checkpoint download size.
Q4 planning estimate only. Check checkpoint format, runtime support, KV cache, context, batch size and offload before choosing hardware.
Evidence level: Generic estimate only; no execution reproduced by this site.
Before downloading
- Check the exact checkpoint and license linked here.
- Confirm a compatible runtime and precision or quantization for your platform.
- Allow for context, batch size and any extra media components beyond the weights.
Catalog reviewed: 2026-09-25. Benchmark release: 2026-06-25; snapshot collected: 2026-09-22.