PyTorch 2.13 Released: FlexAttention on Apple Silicon, 4× Peak Memory Cut for Large Vocab Models

PyTorch · x · 2026-08-01

PyTorch 2.13 has been officially released, bringing several key performance optimizations and feature expansions:

A live Q&A session also covered CUDA version support, CuTeDSL, Python 3.15, and plans for PyTorch 2.14.

Related event: PyTorch 2.13 Released: FlexAttention Hits Apple Silicon(2 posts)→

Original post →

More from Infra

Infra channel →