02 / Memory estimate
Planning formula: total parameters × 0.625 bytes, plus 20% working allowance and 8 GB reserve, rounded up in decimal GB. This is a heuristic, not a measured minimum or guarantee. Quantization, context, KV cache, vision components, runtime and offloading change requirements. MoE uses total parameters, not only active ones.
Qwen counts refer to language parameters; GLM-5.3 and DeepSeek V4 use repository tensor totals; GLM-5.3 Flash uses the manufacturer’s 320B specification. Architectures without a reviewed estimate remain outside the memory axis.
03 / Reference, not a local test
LiveBench evaluates the named model and reasoning effort. Its score is not a measurement of the local Q4 variant. Hardware links are official destinations without affiliate tracking. Stock, exact configuration and regional availability can change.
04 / Automatic updates without paid APIs
The deployed product checks scores, approved official model organizations, US prices and stock daily, and hardware specifications weekly. Due visits can also refresh safely. D1 preserves the last valid snapshot. Incompatible methodology is quarantined; unavailable sources never erase valid data. New models remain pending benchmark or insufficient data until official facts exist.