Inkling · xhigh
Quality, size and local speed
#14 among 22 evaluated open models in this release
Where it performs best
LiveBench 2026-06-25 · 0–100 ↗LiveBench 2026-06-25 measures the named benchmark variant, not a local Q4 run. Scores were collected 2026-09-26; the seven categories share the same 0–100 scale. This is not the Artificial Analysis Intelligence Index. Q4 size covers only the counted weights; multimodal auxiliary components may add files.
Quality versus estimated Q4 weight size
Reviewed language checkpoints with a LiveBench score and a parameter-based size estimate. The highlighted point is this model when both values are available.
Horizontal axis: estimated Q4 model-weight disk size in decimal GB (log scale). Vertical axis: LiveBench overall / 100. This formula is not the published checkpoint size, runtime memory or a speed test. Models without both inputs are left out of this chart, not scored as zero.
The Inkling · xhigh checkpoint is published by Thinking Machines for language and reasoning. Its weights are available under Apache 2.0; that alone does not establish that its code and all components are open source.
Compare Inkling · xhigh with other language and reasoning checkpoints using its linked publisher card and the LiveBench 2026-06-25 snapshot. This is quality evidence, not a local speed result.
The generic Q4 planning estimate is about 740 GB from 975B stated parameters. It is not a measured minimum; context, batch size, runtime and quantization change actual memory use.
What is documented
- Checkpoint / version
- thinkingmachines/Inkling ↗
- Weight license
- Apache 2.0 ↗
- Task and modality
- language and reasoning · language model; inputs not listed
- Count and basis
- 975B · manufacturer specification · 41B active
- Context
- 1,000,000 tokens published for this model
- Quality evidence
- 71.9 · LiveBench 2026-06-25 overall · source ↗
- Runtime and precision
- No local execution verified. Check the publisher's supported runtime and checkpoint precision; batch and context memory are not measured here.
Run locally: known and unknown
~740 GB generic Q4 planning memory. Formula: parameters × 0.625 bytes, plus 20% working allowance and 8 GB reserve. This is neither measured VRAM nor checkpoint download size.
Q4 planning estimate only. Check checkpoint format, runtime support, KV cache, context, batch size and offload before choosing hardware.
Evidence level: Generic estimate only; no execution reproduced by this site.
Before downloading
- Check the exact checkpoint and license linked here.
- Confirm a compatible runtime and precision or quantization for your platform.
- Allow for context, batch size and any extra media components beyond the weights.
Catalog reviewed: 2026-09-30. Benchmark release: 2026-06-25; snapshot collected: 2026-09-26.