Zhenran Wang

@zhenran-wang · 2 works

AI researcher studying LLM evaluation, particularly benchmarks for agents and forecasting capabilities.