DeepSeek V4.1 Flash · max
Quality, size and local speed
#1 among 22 evaluated open models in this release
Where it performs best
LiveBench 2026-06-25 · 0–100 ↗LiveBench 2026-06-25 measures the named benchmark variant, not a local Q4 run. Scores were collected 2026-09-26; the seven categories share the same 0–100 scale. This is not the Artificial Analysis Intelligence Index. Q4 size covers only the counted weights; multimodal auxiliary components may add files.
Quality versus estimated Q4 weight size
Reviewed language checkpoints with a LiveBench score and a parameter-based size estimate. The highlighted point is this model when both values are available.
Show 10 more models
Qwen3.8 27B75.3~16.9 GBInkling · xhigh71.9~609.4 GBGLM-5.3 Flash71.6~200.0 GBDeepSeek V4 Pro71.6~1000.0 GBKimi K2.6 · thinking70.5~625.0 GBKimi K2.7 Code68.4~625.0 GBNemotron 3 Ultra 550B A55B67.4~343.8 GBMiniMax M367.3~267.5 GBDeepSeek V4 Flash65.5~177.5 GBQwen3.6 27B64.0~16.9 GBHorizontal axis: estimated Q4 model-weight disk size in decimal GB (log scale). Vertical axis: LiveBench overall / 100. This formula is not the published checkpoint size, runtime memory or a speed test. Models without both inputs are left out of this chart, not scored as zero.
The DeepSeek V4.1 Flash · max checkpoint is published by DeepSeek for language and reasoning. Its weights are available under MIT; that alone does not establish that its code and all components are open source.
Compare DeepSeek V4.1 Flash · max with other language and reasoning checkpoints using its linked publisher card and the LiveBench 2026-06-25 snapshot. This is quality evidence, not a local speed result.
This profile has no simplified memory estimate because the available size evidence is incomplete or includes separate components. Check the exact checkpoint, format and runtime before planning hardware.
What is documented
- Checkpoint / version
- deepseek-ai/DeepSeek-V4.1-Flash ↗
- Weight license
- MIT ↗
- Task and modality
- language and reasoning · text + image
- Count and basis
- Not verified
- Context
- 1,000,000 tokens published for this model
- Quality evidence
- 81.1 · LiveBench 2026-06-25 overall · source ↗
- Runtime and precision
- No local execution verified. Check the publisher's supported runtime and checkpoint precision; batch and context memory are not measured here.
Run locally: known and unknown
No validated memory estimate for this model and runtime.
Q4 planning estimate only. Check checkpoint format, runtime support, KV cache, context, batch size and offload before choosing hardware.
Evidence level: Generic estimate only; no execution reproduced by this site.
Before downloading
- Check the exact checkpoint and license linked here.
- Confirm a compatible runtime and precision or quantization for your platform.
- Allow for context, batch size and any extra media components beyond the weights.
Catalog reviewed: 2026-09-30. Benchmark release: 2026-06-25; snapshot collected: 2026-09-26.