Polymorf pushes OMLX inference from 150tps to nearly 200tps within 24 hours of launch
HankYeomans · x · 2026-09-26
X user TheDavidTai reports that Polymorf, less than 24 hours after launch, smashed the 166tps OMLX reference benchmark he had been using—boosting inference speed from 150tps to almost 200tps in one step. Technical details were not elaborated in the post.
More from Infra
- SpaceX's Memphis supercomputer: millions of GPUs and over two gigawatts of compute — CurieuxExplorer · 2026-09-26
- The handiest GPU this dev ever bought is a ~$300 Intel Arc A310, not NVIDIA or AMD — TheZachMueller · 2026-09-26
- Running image generation in the browser on local hardware: 10s pixel art on an RTX 3060 — Bartholomheow · 2026-09-26
- Alibaba's T-Head unveils Zhenwu V900 chip: 216GB per card, sales in Q1 2027 — shashib · 2026-09-26
- llama.cpp fork dedups repeated prompts losslessly, cutting 108k to 71k tokens in agent loops — Odd_Cauliflower_8004 · 2026-09-26
- Why rent servers when agents can run your terminal? — StewartalsopIII · 2026-09-26