UkisAI ships Swift reasoning LLM family: -63.4% thinking tokens at 1.8x speed on Qwen base

Secure_Recording_472 · reddit · 2026-09-25

Small lab UkisAI released its Swift family of efficient reasoning LLMs based on Qwen, trained by penalizing pathological overthinking tokens and restoring accuracy via GSPO RL and on-policy distillation. The previous Swift Qwen 3.8 27B hit 350k+ downloads in 13 days.

This release:

Benchmarks run 5x across seeds over GPQA, AIME26, LiveCodeBench, ERQA and Terminal Bench 2.1. GGUF/NVFP4/MLX/W4A16 quants, a Research API and HF Spaces are available; a 9B variant is coming soon.

Original post →

More from Models

Models channel →