CritICL Boosts Big Model Reasoning Using Small Model Failure Modes
rohanpaul_ai · x · 2026-08-31
Paper "CritICL" proposes using small models' mistakes to improve large model reasoning. Since failure modes are structured across model scales, CritICL runs small models on math problems, saves wrong answers with critiques, and retrieves relevant critiques for the big model's prompt during inference. Results show this method achieves accuracy comparable to test-time scaling (e.g., majority voting) with just 1 generation, significantly reducing token costs and compute requirements.
Related event: CritICL Uses Small Model Failures to Boost LLM Reasoning(2 posts)→
More from Research
- LiteMol-1 generates drug candidates on M1 Max in 30 seconds — CatAstro_Piyush · 2026-09-01
- 40-nm Memristor Chip Turns Conductance Drift Into a Feature, Beats A100 by 50-480x — maier_ak · 2026-09-01
- Google Paper: Autonomous AI Research Hallucinates 90% Without Checks — rohanpaul_ai · 2026-09-01
- RLHF impact on tokens: unconscious shifts vs conscious choices — voooooogel · 2026-09-01
- On token layers and consciousness in RLHF — voooooogel · 2026-09-01
- CommerceAgentBench released: Qwen leads open-weight models — Alibaba_Qwen · 2026-09-01