RTX 3090 runs 27B Qwen at 150k context with DeepSeek harness, results amaze

politefella0 · reddit · 2026-08-25

A Reddit user shares hands-on results with syv-ai/qwen38-27b-rtx3090: on a single RTX 3090 with vision enabled, they ran 150k context with excellent results, even having local Qwen write a Gmail plugin for the DeepSeek harness. A search-engine plugin attempt broke dsh and made the harness unlaunchable.

Run stats: 26 turns / 489 steps, 161m LLM time, 10m tool calls, 5.9s avg TTFT, 86 tok/s, 0% cache hit, 35.7M input / 586K output tokens.

Original post →

More from coding & agent

coding & agent channel →