Qwen3.8-27B runs 100k-token physics simulation on a 16GB GPU

1000_bucks_a_month · reddit · 2026-08-26

A developer ran a heavily quantized Qwen3.8-27B model (IQ3XXS) on an older 16GB Quadro RTX 5000, tasking it with implementing the coherent optical transfer-matrix method (TMM) for multilayer films from scratch. The session lasted about 100 minutes, generating 108,000 output tokens and undergoing three context compaction attempts.

Key Details:

Original post →

More from coding & agent

coding & agent channel →