Explain the arrangement-dependent sign reversal in the post-trained Qwen3.5-9B model
Explain why the post-trained Qwen3.5-9B model produces arrangement-dependent sign reversals in the influence of a risk disclosure across document layouts, with negative effects when filler precedes the focal filing and positive effects when filler follows it.
References
This is not an artifact of low capability---the post-trained 9B has the strongest severity discrimination in the entire ladder ($S_7-S_1$ range of $+0.42$, versus $+0.23$ for its base counterpart; Appendix Table~\ref{tab:app_sevgate})---and we do not have a mechanism for it.
— Reading Is Not Using: Retrieval, Judgment, and the Design of AI Financial Research Workflows
(2608.24842 - Liu et al., 25 Aug 2026) in Section 3.5, “Capability Moves the Horizon Outward”