Help: Converting Qwen3.8 27B to ONNX for NPU Usage
xXDennisXx3000 · reddit · 2026-08-17
A user seeks help converting Qwen 3.8 27B from safetensors (bf16) to ONNX format. The goal is to enable full utilization of hybrid NPU + GPU setups on devices like Strix Halo.
The user reports failed attempts over 3-4 days on Linux Ubuntu and requests community assistance to complete the conversion and share the model.
More from Infra
- Data Center Boom Becomes a Texas Story as Demand Surges — bigdata · 2026-08-17
- Squeezing Qwen 3.8 27B MTP Q8_0 + Vision into 48GB VRAM — Creative-Type9411 · 2026-08-17
- AI’s next bottleneck is power: The race for energy — ingliguori · 2026-08-17
- ComfyUI OOMs after 20 generations: MiniMaxH3 suspected memory leak — reicaden · 2026-08-17
- System design in 2026: Master these 10 core concepts — blaizedsouza · 2026-08-17
- AI Economics: Why compute and power are the new bottlenecks — demian_ai · 2026-08-17