Decision 3.0 ported to Core ML: 1,000 PRs triaged on-device at 0.12s each

emax · x · 2026-10-12

The FluidInference community has converted the vLLM Semantic Router team's open-source Decision 3.0 decision models to Core ML, running fully locally on Apple silicon.

Original post →

More from coding & agent

coding & agent channel →