Swift Qwen 3.8 27B hits 100k+ downloads, cutting token use 58.3% with 1.95x speed
Secure_Recording_472 · reddit · 2026-09-18
Jovan from small lab UkisAI thanked the community as Swift Qwen 3.8 27B passed 100k downloads, ranking #1 finetune and #9 overall on HuggingFace Trending.
- The approach penalizes pathological overthinking patterns in small LLMs rather than training them to think shorter: token usage drops 58.3% and speed improves 1.95x with no accuracy loss
- Upcoming releases: Swift1.5 Qwen3.8 27B (fixed training bugs, more RL) and Swift Qwen3.8 Flash Next next week, with an expanded benchmark suite including coding and long-horizon tasks
- The team is soliciting community requests on quants and features for future releases
More from Models
- Qwen 3.8 Omni Flash Surfaces with Continuous Video/Audio Understanding and Custom Harness — Mr_Moonsilver · 2026-09-18
- Encoders and decoders are the same thing, and decoders have been doing classification for years — HanchungLee · 2026-09-18
- Sakana AI Launches Fugu Max, a Multi-Agent Orchestrator Routing Tasks to Leanest Capable Models — tkasasagi · 2026-09-18
- Prediction: Every Future LLM Will Ship With a Native 'Jev Mode' — multimodalart · 2026-09-18
- Why don't LLM labs compete on personality? Users value it over raw capability — dioscuri · 2026-09-18
- Fine-tuned 4B model as a decision scorer with temperature-scaled confidence — Gradio · 2026-09-18