Local LLM on Mac: M2 Ultra 192GB Long-Context Inference Benchmarks

Badger-Purple · reddit · 2026-08-02

A developer tested the local inference performance of a large model (Deepseek-V4-Flash-0731 Dwarfstar) on an M2 Ultra machine with 192GB of unified memory.

Performance Data:

Original post →

More from Infra

Infra channel →