Benchmarks Confirm ThinkingCap & Swift Cut Qwen3.8-27B Reasoning Tokens ~40% With Minimal Loss

returnity · reddit · 2026-09-25

An independent Aider eval suite comparison of ThinkingCap-Qwen3.8-27B, Swift-Qwen3.8-27B, and vanilla Qwen3.8-27B validates both fine-tunes' claims of 40% reasoning token reduction with minimal performance loss.

Key results (2 runs per model, ±2-3% error bars):

Nuances:

Original post →

More from Models

Models channel →