Scaling the quality curve to 70B models

Determine whether the quality curve of Leech-lattice vector quantization, whose direction has been measured on Qwen3-4B, Qwen3-8B, and Qwen3-14B without establishing a scaling law, reaches the 70B model class.

Background

The paper evaluates the same 2-bit Λ24(12)\Lambda_{24}(12) quantization configuration on Qwen3-4B, Qwen3-8B, and Qwen3-14B. Across these three sizes, perplexity degradation and the MMLU deficit decrease with model size, but the authors explicitly decline to infer a scaling law because the evidence comprises only three models and is sensitive to calibration draws.

The 70B class is the regime motivating the paper’s memory argument, but it lies outside the measured model sizes. The paper therefore leaves unresolved whether the observed improvement in quality with increasing model size continues at 70B.

References

Two questions stay open: does the quality curve, whose direction holds on three sizes without a scaling law, reach the 70B class, and which untested lever (calibration composition, learned column scales, low-rank compensation) closes the reasoning-concentrated MMLU deficit?

Unfolding the Leech Lattice: Fused Multi-Shell Decoding and VRAM Layouts for 2-Bit LLM Weights  (2609.02652 - Malandrino, 2 Sep 2026) in Conclusion, final paragraph; see also Section 5, Section 7, and Appendix A