Dev benchmarks Bend 2 on M4 Max: parallel kernel 6.4x faster than NumPy

arthurcolle · x · 2026-09-20

A developer shared first-hand benchmarks of a small F32 state-update kernel written with Bend 2 on an M4 Max:

Original post →

More from coding & agent

coding & agent channel →