Skill Self-Play: LLMs Learn by Co-Evolving Skills, No Humans Needed
Skill Self-Play (SSP) proposes a novel training paradigm where LLMs generate and verify their own tasks through co-evolving skills, addressing the diversity-reliability dilemma. This could reshape how LLMs are improved, but questions about scalability and robustness remain.











