PyTorch MPS Linear Algebra Called 'Trash': Slow and Improperly Batched
ducha_aiki · x · 2026-09-03
- Developer duchaaiki reports that linear algebra ops on PyTorch's MPS (Apple Silicon) backend perform poorly.
- Two issues: operations are very slow, and batching is not properly implemented.
- Useful pitfall intel for anyone training or running inference locally on Apple Silicon.
More from Infra
- Build a Fully On-Device Voice ChatGPT for iOS with Apple's Free Models — amos_gyamfi · 2026-09-03
- Microsoft to Disclose Azure Revenue Quarterly in Major Financial Reporting Overhaul — tomwarren · 2026-09-03
- HyperspaceDB v3.1.4: 1-bit ADC gets 107× search speedup, ships Mem0 drop-in replacement — Sam_YARINK · 2026-09-03
- Tesla targets one Cybercab every 5 seconds from a single production line — XFreeze · 2026-09-03
- New paper debunks RL batch scaling: 2.29x throughput doesn't mean faster learning — teortaxesTex · 2026-09-03
- Anthropic Signs $35 Billion Cloud Deal With Lambda to Scale Claude — The Decoder · 2026-09-03