Qwen3.8-27B benchmark: BF16 TP2 performance data
Maleficent_Bridge_41 · reddit · 2026-08-15
A Reddit user shared benchmark results for Qwen3.8-27B on llm-decode-bench, using BF16 precision and TP2 configuration (2x RTX 6000) with vLLM nightly.
Related event: Qwen3.8-27B Benchmarks and Real-World Tests Draw Attention(2 posts)→
More from Models
- Orion-16B passes 100B tokens, largest LLM pretrained with decentralized compute — const_reborn · 2026-08-15
- AI model benchmarks: baseline crushed, top models nearly finish course — const_reborn · 2026-08-15
- Qwen3.8 gets Day-0 support from LightSeek, boosting inference performance by 30%+ — Alibaba_Qwen · 2026-08-15
- DeepSeek's inference cost advantage? User says GLM 5.3 hit daily limits while DS handled it easily — teortaxesTex · 2026-08-15
- Grok 4.6 Boosts Performance: Fable-Level Intelligence, Sonnet Speed, Lower Cost — vikvang1 · 2026-08-15
- Qwen3.8-27B-FP8 on GH200: 10 concurrent streaming requests, first token in 10ms — MaziyarPanahi · 2026-08-15