Learn meta-skills that co-evolve with reasoning policies through reinforcement learning
Develop reinforcement-learning methods for learning meta-skills that co-evolve with the reasoning policy, rather than relying on predefined meta-skill programs or fixed workflows.
References
Learning meta-skills that co-evolve with the reasoning policy through RL therefore remains open.
— CoSkill: Joint Reinforcement Learning of Reasoning and Meta-Skill Agents for Hierarchical Skill Evolution
(2609.04865 - Feng et al., 4 Sep 2026) in Section 1, Introduction