Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills
This paper develops a new framework called Skill Self-Play that helps large language models (LLMs) improve their capabilities by co-evolving skills that balance task diversity and verification reliability. Practitioners might care because this approach can lead to significant performance gains for LLMs.