Global foundation-model progress briefing
English Edition中文
Enter keywords to search ingested stories.

Today / Sunday, September 27, 2026

limbo logolimbo

Data updated

Aug 27, 05:59 PM

Live sources

17

Ingestion status

Live ingest

ResearcharXiv AI / CL

CritICL: Inference-Time Weak-to-Strong Generalization from Small Language Model Failure Modes

Summary

Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated generation or external verification.

Original Article

Captured source content or English translation, normalized into this reading format.

Read Source
This story does not yet have captured source text. Open the source link to read it.

Region

Global

Heat Score

76

Category

Research

Language

en