TypeSafeAI served a trillion tokens on Modal within four days of Jev launch
josh_wills · x · 2026-10-08
In a case study reshared by Modal, the TypeSafeAI team launched their product Jev on a Thursday and had served a trillion tokens by Sunday, with no signs of slowing down. The post links to an engineering write-up on how they scaled Jev on Modal, showcasing the platform's inference infrastructure under heavy production load.
More from Infra
- Windows demo routes coding tasks to local model with GPU spinning, llama.cpp lands on Windows ML — ryanshrout · 2026-10-08
- exe.dev explains crossing the hyper-thread boundary: core scheduling cookies for VM isolation — davidcrawshaw · 2026-10-08
- NVIDIA's LoGRA cuts RL training memory by up to 45.7%, trains 27B model where Adam OOMs — mark_k · 2026-10-08
- Qwen3.8-Flash-Next on 6x3090 without NVLink: prefill 8-10x faster, long-context decode 2-3x — flynth92 · 2026-10-08
- Google DeepMind's Philipp Schmid: Give Every AI Agent Its Own Managed Cloud Sandbox — AI Engineer · 2026-10-08
- GitHub goes down; engineers confirm a fix is in progress — iannuttall · 2026-10-08