Help: DSH model running out of output tokens
Ed-2-Zero-9 · reddit · 2026-08-27
A developer running the DSH model via Unsloth is encountering an 'out of output tokens' issue, despite maxing out the settings in both DSH and Unsloth. The user asks whether to continue tweaking the Unsloth setup or switch to a different inference engine that supports load balancing for gfx120x architecture (R9700 and 9070 XT).
More from Infra
- Nvidia projects 70% revenue growth for fiscal 2028, beating analyst expectations of 44% — firstadopter · 2026-08-27
- Nvidia's NVHBM Brings 30% More Bandwidth to NVLink Fusion; Amazon Annapurna First Partner — nordicinst · 2026-08-27
- NVIDIA CFO forecasts 70% growth next year, targeting ~$700B revenue — BenBajarin · 2026-08-27
- AWS and NVIDIA to Deploy 2 Million Additional GPUs for Agentic and Physical AI — nvidia · 2026-08-27
- Gemini 3.5 Transcribe Now Available on Vercel AI Gateway — osanseviero · 2026-08-27
- Nvidia guides $108B Q3 revenue, doubling growth even without China data center sales — inductionheads · 2026-08-27