Qwen3-Coder 30B runs locally in GitHub Copilot at ~85 tokens/s on 96GB VRAM

ollama · x · 2026-08-13

Developer burkeholland ran Qwen3-Coder 30B locally in the GitHub Copilot app via Ollama, achieving 85 tokens per second on 96GB VRAM. Not as fast as Sol, but promising progress.

Original post →

More from coding & agent

coding & agent channel →