PaperGym: Rubric-Centered Evolution for Research-Plan Generation
By Yuhan Wang · Paper · cs.CL
Research planning is the decisive capability of AI scientists. Yet a research plan admits no verifiable answer, so reinforcement learning lacks the environment it requires: tasks paired with a critic. Rubrics extracted from scientific papers can supply the critic. Existing pipeli