1.7B specialized model outperforms Qwen3-8B on strict logic benchmarks

Creative-Fig522 · reddit · 2026-08-17

TwIL-LM2, a LoRA adapter on SmolLM2-1.7B, achieves impressive results on strict formal logic translation tasks. Its Strict-7 score (0.2386) surpasses Qwen3-8B (0.2093) and Gemma-4-26B (0.2050). This demonstrates that small models, through deep specialization, can outperform much larger LLMs in specific, narrow reasoning tasks, challenging the narrative that complex reasoning strictly requires massive scale.

Related event: 1.7B Logic-Tuned Model Outperforms Qwen3-8B on Formal Logic(2 posts)→

Original post →

More from Models

Models channel →