Identify the training mechanisms producing scaled-idempotent orientation
Identify the data, gradients, and downstream objectives that produce the trained transport orientations responsible for the sparse attainment of high scaled idempotence in Transformer attention OV operators.
References
Our experiments characterize that separation geometrically; identifying the data, gradients, and downstream objectives that produce it remains an open question.
— Scaled Idempotence in Transformer Attention: Paired OV Geometry and Shared-Value Algebras
(2609.01129 - Feng et al., 1 Sep 2026) in Section Discussion, subsection “What $T^2\approx\alpha T$ encodes”