Ling 3.0 Flash
Quality, size and local speed
No comparable score in this release
LiveBench 2026-06-25 measures the named benchmark variant, not a local Q4 run. Scores were collected 2026-09-26; the seven categories share the same 0–100 scale. This is not the Artificial Analysis Intelligence Index. Q4 size covers only the counted weights; multimodal auxiliary components may add files.
Quality versus estimated Q4 weight size
Reviewed language checkpoints with a LiveBench score and a parameter-based size estimate. The highlighted point is this model when both values are available.
Show 10 more models
Qwen3.8 27B75.3~16.9 GBInkling · xhigh71.9~609.4 GBGLM-5.3 Flash71.6~200.0 GBDeepSeek V4 Pro71.6~1000.0 GBKimi K2.6 · thinking70.5~625.0 GBKimi K2.7 Code68.4~625.0 GBNemotron 3 Ultra 550B A55B67.4~343.8 GBMiniMax M367.3~267.5 GBDeepSeek V4 Flash65.5~177.5 GBQwen3.6 27B64.0~16.9 GBHorizontal axis: estimated Q4 model-weight disk size in decimal GB (log scale). Vertical axis: LiveBench overall / 100. This formula is not the published checkpoint size, runtime memory or a speed test. Models without both inputs are left out of this chart, not scored as zero.
The Ling 3.0 Flash checkpoint is published by InclusionAI for language and reasoning. Its weights are available under MIT; that alone does not establish that its code and all components are open source.
Compare Ling 3.0 Flash with other language and reasoning checkpoints using its linked publisher card. No score for it appears in the reviewed LiveBench snapshot. This is quality evidence, not a local speed result.
The generic Q4 planning estimate is about 101 GB from 124B stated parameters. It is not a measured minimum; context, batch size, runtime and quantization change actual memory use.
What is documented
- Checkpoint / version
- inclusionAI/Ling-3.0-flash ↗
- Weight license
- MIT ↗
- Task and modality
- language and reasoning · language model; inputs not listed
- Count and basis
- 124B · manufacturer specification · 5.1B active
- Context
- 262,144 tokens published for this model
- Quality evidence
- No comparable score in the reviewed snapshot
- Runtime and precision
- No local execution verified. Check the publisher's supported runtime and checkpoint precision; batch and context memory are not measured here.
Run locally: known and unknown
~101 GB generic Q4 planning memory. Formula: parameters × 0.625 bytes, plus 20% working allowance and 8 GB reserve. This is neither measured VRAM nor checkpoint download size.
Q4 planning estimate only. Check checkpoint format, runtime support, KV cache, context, batch size and offload before choosing hardware.
Evidence level: Generic estimate only; no execution reproduced by this site.
Before downloading
- Check the exact checkpoint and license linked here.
- Confirm a compatible runtime and precision or quantization for your platform.
- Allow for context, batch size and any extra media components beyond the weights.
Catalog reviewed: 2026-09-30. Benchmark release: 2026-06-25; snapshot collected: 2026-09-26.