Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills
Siyuan Huang, Pengyu Cheng, Haotian Liu, et al.
Skill Self-Play (Skill-SP) co-evolves a proposer, solver, and skill controller via RL to reconcile structured verification with open-ended task diversity for LLM self-evolution.