Skill Self-Play: New Co-Evolution for LLM Training Methods27. July 2026AI ModelsSkill Self-Play combines task generation, solution search, and dynamic skill control in a reinforcement learning loop to achieve both task diversity and training reliability. Share on: