Qwen 3.8 runs locally for 1.7 hours to generate AI daily news

Teknium · x · 2026-08-16

A user ran Qwen 3.8 (27B) locally on a Mac Mini with 24GB RAM as the top orchestrator within the Hermes Agent framework. The agent successfully aggregated the last 24 hours of AI news and generated a voice-over in a specific style, taking approximately 1 hour and 42 minutes. Benchmark speeds were around 16.8 tokens/sec for generation and 40.6 tokens/sec for prompt evaluation. The user plans to migrate the task to an RTX-5090 setup for significantly faster performance.

Original post →

More from coding & agent

coding & agent channel →