Determine whether expert routing distributes compliance computation
Determine whether expert routing in sparse mixture-of-experts language models distributes prompt-injection compliance computation rather than concentrating it in a localized single-token causal bottleneck.
References
We therefore make no strong claim that the dense single-token bottleneck carries over to sparse MoE routing, and whether expert routing distributes the compliance computation is an open question.
— Where Do LLMs Decide to Break the Rules? Mechanistic Localization of Prompt Injection Compliance
(2609.37737 - Wen et al., 29 Sep 2026) in Appendix, Section "Additional Models: Reasoning, MoE, and Training Stage," subsection "Reasoning and Mixture-of-Experts" (Appendix A.6.1)