Benchmarks of models on Radeon 680M iGPU

tabletuser_blogspot · reddit · 2026-08-17

Benchmarked 9 models (including Bonsai-27B, Gemma-4-12B) on a mini PC with AMD Radeon 680M iGPU. Using llama.cpp's Vulkan backend and FlashAttention, reported prefill and generation speeds, showing smaller models achieve high throughput on iGPUs.

Original post →

More from Infra

Infra channel →