ResearcharXiv AI / CL
CritICL: Inference-Time Weak-to-Strong Generalization from Small Language Model Failure Modes

Summary
Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated generation or external verification.
Original Article
Captured source content or English translation, normalized into this reading format.
This story does not yet have captured source text. Open the source link to read it.
Region
Global
Heat Score
76
Category
Research
Language
en
