Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

By Siyuan Huang · Paper · cs.CL

LLM training is shifting from manual design and annotation to interaction-driven self-evolution. However, existing self-evolutionary methods face a fundamental dilemma between task diversity and verification reliability: environment-bound methods obtain precise feedback but confi

Cs.cl

View original

HomeResourceLoading…