Xingyu Shen

@xingyu-shen · 1 works

Researcher working on reinforcement learning methods to improve LLM reasoning and policy entropy.