Swift-Qwen3.8-27b, a token-efficient reasoning Qwen finetune, trends on Hugging Face
ukisai · hf · 2026-09-14
The community model ukisai/Swift-Qwen3.8-27b is trending on Hugging Face, built on the Qwen3.5/Qwen3.8 architecture with an image-text-to-text pipeline.
- Tagged efficient-thinking, token-efficient, and reasoning, emphasizing token-efficient inference
- Trained with LoRA, released in safetensors format for transformers-based conversational use
More from Models
- David Bellamy clarifies his experiment used K2 Horizon, an open-weights 375B LLM — JeremyNguyenPhD · 2026-09-14
- GPT-6 Astra hands-on: composes first, orchestrates later, and reportedly outshines Fable and Sol — paw_lean · 2026-09-14
- Researcher posts proof he both synthesized viruses and trained a 375B open-weight LLM — ethanCaballero · 2026-09-14
- Toby Ord: 10x more RLVR compute cuts tokens-to-target ~3x; gains may be math-specific — tobyordoxford · 2026-09-14
- New scaling curve has half the slope: 10,000x compute for 20%-to-80%, but bigger generational jumps — tobyordoxford · 2026-09-14
- Fudan NLP paper explains why max reasoning settings can backfire on SWE benchmarks — karminski3 · 2026-09-14