1.7B Model Outperforms Qwen3-8B in Strict Formal Logic Reasoning
Creative-Fig522 · reddit · 2026-08-22
A Reddit user highlighted TwIL-LM2, a 1.7B parameter model that achieved surprising results in strict formal logic reasoning, beating larger competitors like Qwen3-8B and Gemma-4-26B.
- Specs: TwIL-LM2 is a PEFT LoRA adapter based on SmolLM2-1.7B, specialized purely for formal logic translation.
- Benchmarks: On the strict-7 scoring (no partial credit, exact format required), it scored 0.2386, ahead of Qwen3-8B (0.2093) and Gemma-4-26B (0.2050). However, larger models still win on the loose-match average.
- Implication: This challenges the narrative that reasoning strictly requires massive scale. If high efficiency can be achieved via narrow specialization on small models (1-3B), it suggests a future pipeline of specialists rather than routing everything to a monolithic 70B model.
- License: Non-commercial license.
More from Models
- DeepSeek launches V4-Flash-Vision-Exp, multimodal agent performance nears Opus-4.8 — OwariDa · 2026-08-22
- EMNLP 2026 Paper: RAG over Thinking Traces Boosts Reasoning by 43% — matei_zaharia · 2026-08-22
- Qwen-3.8 27B Uncensored Model Tested: Discussing Cortés Without Moral Lectures — doodlestein · 2026-08-22
- Ling-3.0-flash-dspark open-sourced, hits 1,120 tok/s on 4 Blackwell GPUs — AdinaYakup · 2026-08-22
- llama.cpp Adds Support for 280B Parameter Multimodal Model dots3-note — jacek2023 · 2026-08-22
- Laurence Moroney on 2026 On-Device Small AI: Gemma 4 & Qwen 3.5 Top Picks — lmoroney · 2026-08-22