CoreML runs FP8-weight model on M6 chip in early patch test
AIFlow_ML · x · 2026-09-24
Developer @anemll shipped a patched CoreML model that loads FP8 weights on Apple's M6 chip, verified working on small tests. Larger model conversion is still pending, marking an early experiment in on-device FP8 deployment.
More from Infra
- Industry responds to hyperscale RDMA paper with Multipath Reliable Connection on path to Ultra Ethernet — thoefler · 2026-09-24
- Amazon nearly doubles US data center capacity since ChatGPT launch, 69% ahead of Microsoft — anselm · 2026-09-24
- JVM runs NVIDIA Parakeet ASR at up to 2X the C++ reference speed, no GPU or Python needed — mkheck · 2026-09-24
- Run a Claude Code-like coding workflow for free locally with Ollama — Aiden_Tech_Ai · 2026-09-24
- NVIDIA researchers show KV caches transfer between models via closed-form linear mapping — anselm · 2026-09-24
- _xjdr: TPUs performing well wasn't a given — time to look at v8s — _xjdr · 2026-09-24