CritICL: Inference-Time Weak-to-Strong Generalization from Small Language Model Failure Modes

By Yufan Wu · Paper · cs.CL

Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated generation or external verification. To address this limitation, we introduce CritICL, a novel infere

Cs.cl

View original

HomeResourceLoading…