Desktop iGPU beats dual 3090s due to a Jinja template

TrifleHopeful5418 · reddit · 2026-08-25

A comparison on LiveCodeBench v6 shows Ornith-1.5-35B-A3B running on a Strix Halo iGPU nearly matches Qwen3.8-27B on dual 3090s. The key finding is that adding a LoRA and switching to a Sharp Chat Template improved scores by 15 problems, a gain larger than adding a second GPU. The author also warns that the default thinking-mode profile in GGUF failed all tests, recommending the Instruct profile instead.

Original post →

More from Models

Models channel →