Tight Characterizations Beyond the Established Scaling Regimes
Establish tight characterizations for data-limited mixed-training regimes and for two-stage training, extending beyond the optimization-saturated mixed-training regime and the upper-bound-only result currently available for two-stage training.
References
In addition, the mixed-training scaling law is tight only in the optimization-saturated regime, while the two-stage result remains an upper bound; this disparity between the two protocols leaves tight characterizations for data-limited and two-stage settings open.
— Learning with Synthetic Data via SGD in High-Dimensional Linear Regression
(2609.09572 - Li et al., 9 Sep 2026) in Section 6, Conclusion, paragraph “Limitations”