Universality and baseline requirements of token-readable contrastive components
Determine whether every model computation yields a token-readable component under contrastive subtraction, and establish how many baselines are sufficient in general for multi-contrast triangulation.
References
We do not know whether every model computation yields a token-readable component under contrastive subtraction, or how many baselines are sufficient in general.
— Contrastive Projection: Reading Transformer Internals by Differencing Logit Lenses
(2609.09902 - Tuomi, 9 Sep 2026) in Section 6.5, Limitations, bullet “Triangulation coverage”