Training stochasticity as a source of quantization-error counteraction
Determine whether training stochasticity produces the counteraction between accumulated block-input hidden error and newly introduced block-update error observed in pretrained Transformer language models.
References
Two questions remain open: whether training stochasticity produces counteraction, as suggested by implicit-bias analyses of SGD~\citep{pmlr-v119-smith20a}, and whether these findings can improve post-training quantization.
— Why Does Post-Training Quantization Work?
(2609.11716 - Chen et al., 10 Sep 2026) in Section 5.4, “Limitations” (Section \ref{sec:scope-limitations})