Taalas 14,000 tokens/sec demo comes from The Peter McCormack Show

iamfakhrealam · x · 2026-08-31

Follow-up: the widely shared Taalas inference demo (100 tokens/sec for ChatGPT vs 14,000 tokens/sec for Taalas) originates from the YouTube channel The Peter McCormack Show, where Emad Mostaque demonstrated the model baked directly into silicon live on air.

Related event: Taalas demos 14,000 tokens/sec inference, ~100x faster than ChatGPT(3 posts)→

Original post →

More from Infra

Infra channel →