Message Passing Enables Efficient Reasoning
By Xuecheng Liu · Paper · cs.CL
While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is a computational bottleneck. Thus, in contrast to sequential scaling methods like CoT, recent parallel scaling techniques instea