Tenstorrent benchmarks for Qwen3.7-27B reveal performance on dual p300c setup
DustNearby2848 · reddit · 2026-08-30
A Reddit user shared benchmarks for running the Qwen3.7-27B model on Tenstorrent hardware. The setup features 2 p300c cards (equivalent to 4 p150s) with 64GB of GDDR6 memory. The benchmarks indicate that the current tests do not yet support MTP (Multi-Token Prediction), providing the community with data to evaluate Tenstorrent's actual LLM inference performance.
More from Infra
- Local LTX 2.5 Generation: 12 Minutes for 20 Seconds of Video? — huynguyend · 2026-08-30
- Bot Mesh: A social network with identity and payments for AI agents — Daniel_Farinax · 2026-08-30
- User Switches to Local Qwen 3.8 27B for Coding to Save API Costs — 4310sy · 2026-08-30
- Bezalel Offers Integrated Super Powers for AI Agents — Rasmic · 2026-08-30
- 19 General Latency Optimization Patterns for Faster AI Applications — blaizedsouza · 2026-08-30
- Superwall's side project policy leads to creation of open-source observability platform Maple — JordanMorgan10 · 2026-08-30