Best Qwen 3.8 Variant for RTX A2000 with 12GB VRAM?
MrMrsPotts · reddit · 2026-08-22
A user with an RTX A2000 (12GB VRAM, 32GB RAM) is currently running qwen3.8-27b-ud-q4km at about 6 tokens/sec. Overwhelmed by the numerous quantization options available, they are seeking recommendations for a configuration that offers the best possible coding quality at a decent speed.
More from coding & agent
- Claude Code now supports starting remote sessions from mobile — majidmanzarpour · 2026-08-22
- Crafting 1,000-line Bash installers quickly with a Claude Code Skill — doodlestein · 2026-08-22
- OpenKnowledge releases open-source plugin for Google's OKF LLM wiki standard — solyarisoftware · 2026-08-22
- JackRabbitOS Released: Open Source Voice Platform for Rabbit R1 — jesselyu · 2026-08-22
- AI Agent Payment Risks: How to Restrict Spending to Approved Merchants? — Hopeful-Horse7580 · 2026-08-22
- Study: 25% of Public MCP Configs Leak API Keys in Plaintext — checkpointdev · 2026-08-22