TwIL-LM2 1.7B outperforms larger models in strict formal reasoning

HopefulMode429 · reddit · 2026-08-23

TwIL-LM2, a 72M param LoRA adapter on SmolLM2-1.7B, specializes in converting English to verifiable first-order logic. It achieves a score of 0.2386 on strict-7 (no partial credit), beating Qwen3-8B (0.2093) and Gemma-4-26B (0.2050). This demonstrates the value of specialized small models for precise tasks (like formal translation) within pipelines, complementing larger general models. Licensed under webAI Non-Commercial License.

Related event: Tiny 1.7B TwIL-LM2 Beats Larger Models at Formal Logic Reasoning(2 posts)→

Original post →

More from Models

Models channel →