Run 35B Models on 16GB Machines! QuarkStar Engine Hits Over 80 tok/s

Nicolodeva · reddit · 2026-08-04

A developer has released QuarkStar, a lightweight native inference engine inspired by DwarfStar, focusing on running LLMs on consumer-grade hardware.

The project aims to push the limits of low-end hardware, allowing users who can't afford $3,000+ AI rigs to run large models locally.

Original post →

More from Infra

Infra channel →