Testing Qwen3.8 Vision: SVG Reconstruction & Anti-Benchmaxxing
bonobomaster · reddit · 2026-08-26
The author designed a benchmark for Qwen3.8-27B to test SVG and vision capabilities in a way that resists 'benchmaxxing'. The task is simple: recreate any given photo as an SVG using the prompt "Recreate as SVG".
Extensive testing revealed that the best results come from --image-min-tokens 1024, --reasoning-effort xhigh, --temperature 1.0, and bf16 KV caches. Notably, q40 quantized caches severely degraded output quality, while q80 performed well. The author also suspects the chat template influences output quality.
More from Models
- Tip: opencode's new model is reportedly a nerfed multimodal GLM-5.3 — PawelHuryn · 2026-08-26
- Model Becomes Fastest Trending in Hugging Face History in 40 Minutes — MaziyarPanahi · 2026-08-26
- Qwen 3.8 Flash beats Opus on SWE-bench Pro — Hesamation · 2026-08-26
- Qwen3.8-Flash-Next: New Architecture Targets Ultimate Cost-Efficiency — tosh · 2026-08-26
- Minecraft clone fully vibecoded with local Qwen3.8-27b — liright · 2026-08-26
- ChatGPT 5.6 Sol Still Hallucinates: 3 Real-World Failure Cases — kaljakin · 2026-08-26