UkisAI ships Swift reasoning LLM family: -63.4% thinking tokens at 1.8x speed on Qwen base
Secure_Recording_472 · reddit · 2026-09-25
Small lab UkisAI released its Swift family of efficient reasoning LLMs based on Qwen, trained by penalizing pathological overthinking tokens and restoring accuracy via GSPO RL and on-policy distillation. The previous Swift Qwen 3.8 27B hit 350k+ downloads in 13 days.
This release:
- Swift1.5 27B: -58.5% thinking tokens, +0.35% average score, beats base on Terminal Bench 2.1 by avoiding overthinking error loops
- Swift Flash Next: -63.4% thinking tokens, 1.8x speedup, only -0.2% vs base on xhigh
- Swift Bonsai 2 (experimental): -39.8% tokens, +0.19% score
Benchmarks run 5x across seeds over GPQA, AIME26, LiveCodeBench, ERQA and Terminal Bench 2.1. GGUF/NVFP4/MLX/W4A16 quants, a Research API and HF Spaces are available; a 9B variant is coming soon.
More from Models
- Relace Is Now the Cheapest DeepSeek v4.1 Flash Provider on OpenRouter — ilyasu · 2026-09-25
- Jev matches year-old top models on global geographic understanding, maps extracted — zetalyrae · 2026-09-25
- Claude Opus 5.5 tops SimpleBench with 88.4% score — Profanion · 2026-09-25
- Anthropic resumes billing for safety-blocked requests; 99.7% of users unaffected — ClaudeDevs · 2026-09-25
- Developer claims: nothing holds back Claude models like Claude Code itself — tokenbender · 2026-09-25
- Blogger: Opus 5.5's strength suggests xAI's rumored Astra is smaller than believed — scaling01 · 2026-09-25