Splash 1.3.0 cuts local agent first-token latency from 19s to 1s via SSD offloading on M6 Mac

BeidiChen · x · 2026-10-09

Splash 1.3.0 targets first-token latency for local agents: with SSD offloading enabled, returning to a project drops time-to-first-token from 19 seconds to 1 second, and new tasks start 1.5× faster thanks to the Neural Engine — all on a 24 GB M6 Mac.

Related event: Splash 1.3.0 cuts local agent first-token latency from 19s to 1s via SSD offloading(2 posts)→

Original post →

More from coding & agent

coding & agent channel →