Laya-MLX: open typed decision model runs locally at 7–14 ms per decision, zero tokens
JiliJeanlouis · x · 2026-09-20
- Developer mizorewww ported the open-source typed-decision system Laya to Apple MLX as laya-mlx (576 GitHub stars), with no PyTorch, Transformers runtime, or cloud API.
- Benchmarks: 13.4 ms median end-to-end for a short English typed decision on M3 Max, 7.4 ms with the multilingual checkpoint, 0 output tokens, 1 GB max memory.
- A demo shows the model playing Snake locally at roughly 60 decisions per second, with a cycle safety layer to correct unsafe proposals.
- Weights are on Hugging Face (aac6fef/laya-mlx); quick start via pip install laya-mlx.
More from coding & agent
- SiftRank adds Jev as an LLM ranker to sift needles from data haystacks — dyn___ · 2026-09-20
- TradingAgents: an open-source multi-agent LLM trading framework in Python — mdancho84 · 2026-09-20
- This guy used an AI agent to profile every eligible bachelor in the city for two cents — gregmushen · 2026-09-20
- OpenHarness: open-source workbench for orchestrating coding agents beyond code — dee_hw · 2026-09-20
- Lessons from a cited paper-writing LangGraph agent: token blowups, fake sources, and four fixes — Altruistic-Video-849 · 2026-09-20
- Plugin4Shell zero-click RCE hits Claude Code, Codex, Copilot and Gemini CLI days before NIST IR 8587, exposing the gap in agent authorization — docybo · 2026-09-20