TwIL-LM2 1.7B outperforms larger models in strict formal reasoning
HopefulMode429 · reddit · 2026-08-23
TwIL-LM2, a 72M param LoRA adapter on SmolLM2-1.7B, specializes in converting English to verifiable first-order logic. It achieves a score of 0.2386 on strict-7 (no partial credit), beating Qwen3-8B (0.2093) and Gemma-4-26B (0.2050). This demonstrates the value of specialized small models for precise tasks (like formal translation) within pipelines, complementing larger general models. Licensed under webAI Non-Commercial License.
Related event: Tiny 1.7B TwIL-LM2 Beats Larger Models at Formal Logic Reasoning(2 posts)→
More from Models
- Hyped Gemini "Autistic" Stealth Model Flops in Rigorous Tests — IndraVahan · 2026-08-23
- Why o3 likely isn't from Cursor: Technical analysis against the GLM theory — teortaxesTex · 2026-08-23
- User claims OpenAI is actively sabotaging Codex coding experience — Revolutionalredstone · 2026-08-23
- Dev Review: Gemini 3.7 Flash performs well for simple/medium tasks — brandon_galang · 2026-08-23
- Opinion: Gemini 3.7 Flash barely keeps Google relevant in AI race — ishuagra02 · 2026-08-23
- Qwen3.5 27B GGUF Quantization Launched for 12GB VRAM — soyaakinohara · 2026-08-23