Identify the feature driving attribution-pattern separation
Identify which feature—adjacency topology, the word class of the intervening word, or position—drives the binary switch in attribution patterns within the determiner–noun agreement construction.
References
Our two isolation experiments point in opposite directions (a random-word insertion between determiner and noun does not collapse layer 0, while an adjective interruption does), and we register this contradiction as an open problem rather than resolving it by fiat.
— Pre-carved Niches: The Formation Dynamics of Modular Task Partitions in Early LLM Training
(2609.01170 - Li et al., 1 Sep 2026) in Section 4.3, “A construction-level separation factor, resolved to features but not to mechanism”