Benchmark Request: Qwen3.8-27B Full Precision & FP8 on RTX 6000 Pro with vLLM
HumanDrone8721 · reddit · 2026-08-15
Reddit user HumanDrone8721 calls for benchmark results of Qwen3.8-27B on RTX 6000 Pro with vLLM in full precision and FP8, providing llama-benchy tool and Qwen recipe links. Results to be posted within an hour.
Related event: Qwen3.8-27B Benchmarks and Real-World Tests Draw Attention(2 posts)→
More from Models
- Orion-16B passes 100B tokens, largest LLM pretrained with decentralized compute — const_reborn · 2026-08-15
- AI model benchmarks: baseline crushed, top models nearly finish course — const_reborn · 2026-08-15
- Qwen3.8 gets Day-0 support from LightSeek, boosting inference performance by 30%+ — Alibaba_Qwen · 2026-08-15
- DeepSeek's inference cost advantage? User says GLM 5.3 hit daily limits while DS handled it easily — teortaxesTex · 2026-08-15
- Grok 4.6 Boosts Performance: Fable-Level Intelligence, Sonnet Speed, Lower Cost — vikvang1 · 2026-08-15
- Qwen3.8-27B-FP8 on GH200: 10 concurrent streaming requests, first token in 10ms — MaziyarPanahi · 2026-08-15