RTX 3090 runs 27B Qwen at 150k context with DeepSeek harness, results amaze
politefella0 · reddit · 2026-08-25
A Reddit user shares hands-on results with syv-ai/qwen38-27b-rtx3090: on a single RTX 3090 with vision enabled, they ran 150k context with excellent results, even having local Qwen write a Gmail plugin for the DeepSeek harness. A search-engine plugin attempt broke dsh and made the harness unlaunchable.
Run stats: 26 turns / 489 steps, 161m LLM time, 10m tool calls, 5.9s avg TTFT, 86 tok/s, 0% cache hit, 35.7M input / 586K output tokens.
More from coding & agent
- Grok Build Adds 'Browser Use' Plugin for Local Chrome and Cloud Browsing — elonmusk · 2026-08-25
- OpenAI Live Demo: Hands-Free Workflows Using Voice Agent in Codex — OpenAIDevs · 2026-08-25
- Grok Review: Independent VMs and Browser Access Enable Seamless Multi-Bot Workflows — brandon_galang · 2026-08-25
- Survey on Terminal Agents: Definitions and Evaluation Frameworks — omarsar0 · 2026-08-25
- Weighted Memory Tree: +10 Points on GAIA-Text With 33% Fewer Tokens — dair_ai · 2026-08-25
- xAI launches Grok Bot, an AI teammate that signs into tools, starting at $60 — NicoVerderosa · 2026-08-25