Qwen 3.8 27B scores 34 (99th percentile) on ACT run locally with vision
on_line187 · reddit · 2026-08-20
A Reddit user had Qwen 3.8 27B Instruct (Q80 GGUF, LM Studio, 2× RTX 3090 full offload, 32k context) take two official ACT practice tests — 342 questions fed as raw PDFs to test vision, with no human help and no retries.
Results: composite scores of 36 (the maximum) and 34, overall accuracy 95.3% (326/342), zero blanks. Reading was perfect on both papers (72/72); math near-perfect; English and science 33-35. A 34 is roughly the 99th percentile.
It took 177 minutes for both tests (88 min each) versus 165 minutes allowed per test for humans — slower than expected because the model read questions visually from PDFs rather than plain text.
More from Models
- Dev praises Qwen 3.8 27B: Benchmarked to actually work — rudrank · 2026-08-20
- Zhipu GLM-5.3 scores 69 on official DeepSWE leaderboard — AccBalanced · 2026-08-20
- Zhipu GLM-5.3 DeepSWE Score Jumps to 69 via Post-Training Scaling — haider1 · 2026-08-20
- Poll gauges user retention after DeepSeek price hike — oran_ge · 2026-08-20
- Cohere study finds 14B models outperform larger ones in linguistic reasoning challenge — Cohere_Labs · 2026-08-20
- Train a 135M parameter LM in 2.5 hours: A full journey — EAccelerate_42 · 2026-08-20