fal engineering head: we'll never pre-train, inference compute is the real moat

jfischoff · x · 2026-09-17

fal's head of engineering explains why the generative media cloud will never pre-train a model: the gaps today are in real-time video generation — which fal claims to have cracked with H3 Max — not pre-training, and "nobody in the world has enough inference compute," so fal is aggressively building and procuring compute. The reposter adds that capabilities really emerge in post-training, where they doubled down last year and saw it pay off.

Original post →

More from Companies & People

Companies & People channel →