Scaling of WhiteMatter quality and systems trade-offs
Determine how WhiteMatter’s language-modeling quality and systems trade-offs scale with model size and data, including through evaluations of larger models and optimized end-to-end decoding.
References
These experiments therefore do not establish how the quality or systems trade-offs scale with model size and data. We report cache size and schedule convergence, but do not provide an optimized end-to-end decoding benchmark. Evaluating larger models and optimized end-to-end decoding remains future work.
— WhiteMatter: All-to-All Cross-Layer Connections via KV Mixing
(2608.18486 - Zhang et al., 19 Aug 2026) in Section “Limitations,” paragraph “Empirical scope”