Sharper stability assumptions for classification test-error estimation
Identify sharper assumptions on the training algorithm, weaker than uniform stability, under which the leave-a-window-out estimator consistently estimates classification test error for dependent data.
References
It remains open to provide sharper assumptions on the training algorithm under which our estimator succeeds.
— Next-token functional estimation
(2609.19529 - Nakul et al., 17 Sep 2026) in Section 7, Discussion