AMD Acquires Taalas to Hard-Code AI Models Directly Into Silicon

The Decoder · rss · 2026-08-08

AMD has acquired Canadian startup Taalas, which specializes in hard-coding AI model weights directly into inference chips.

While this approach locks each chip to a single model, it delivers extreme performance. A demo chip achieved over 16,000 tokens per second per user running Llama 3.1-8B. Google is also reportedly developing a similar hardware approach for Gemini.

Original post →

More from Venture

Venture channel →