Stable Domain Adaptation for Stronger Base Models
Determine how to obtain stable performance improvements when adapting the Qwen3.5-9B base model to the dedicated role-playing task, despite its already high initial character-consistency performance and the limitations of the current training data and pipeline.
References
We conducted identical experiments on the Qwen3.5-9B model but did not observe the same significant performance gains as with the Qwen3 series. The core reason is that the Qwen3.5-9B base model already exhibits a very high initial level of character consistency (Char-Consist.). Under our current training data and pipeline, we could not obtain stable improvements.
— KuaiRP Series Role-playing Models Technical Report
(2609.11127 - Wang et al., 10 Sep 2026) in Section Conclusion and Future Work