DeepSeek's 1/30 Cost Training Breakdown

wordgrammer · x · 2026-08-27

The author shares that they spent the day analyzing DeepSeek's papers to understand how the model was trained at a fraction of the cost (reportedly 1/30th). This post serves as an entry point to a detailed technical breakdown (in the quoted tweet) regarding training efficiency and data center economics.

Original post →

More from Infra

Infra channel →