Scaling ethical-preference behavior to larger language models
Determine whether the ethical-preference learning and task-vector behavior observed in Llama-3.2 1B and 3B checkpoints scales to larger language models.
References
Fourth, our experiments are limited to Llama-3.2 1B and 3B checkpoints, and it remains open whether the same behavior scales to larger models.
— Geometry of Values: Task Vector Composition for Ethical Preference Alignment in Language Models
(2609.21094 - Agarwal et al., 17 Sep 2026) in Section 6, Limitations