Establish whether optimizing competence scores induces rigid reasoning
Establish whether optimizing the aggregate competence score defined from verdict stability, monotonicity, decisiveness, and Pareto viability causes language-model agents to adopt rigid rule-following with explicit exceptions rather than context-sensitive adjustments.
References
Whether this materializes under optimization pressure is an empirical question we do not resolve here.
— Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for Coherent Alignment
(2609.05036 - Libert et al., 4 Sep 2026) in Section 3, Limitations