Sentdex benchmarks openjev: 169ms on Dell GB10 vs 137ms on RTX 3090
Sentdex · x · 2026-09-22
Sentdex shares first-hand openjev testing results, finding it compares very strongly to Jev for LLM augmentation.
- Response latency: 169ms on Dell GB10 vs 137ms on an RTX 3090
- He may keep openjev permanently on the GB10s and run multiple Jev servers per device to keep latency low for parallel queries
A rare hands-on latency data point for local LLM deployment. Maxime Abel
More from Infra
- COLIBRI: pure-C zero-dep engine streams 2.8T-param MoE models from disk on consumer hardware — bibryam · 2026-09-22
- PyroDash cuts inference cost 96% by having a 4B model call the big one only when needed — jiqizhixin · 2026-09-22
- Underdog's Husky Inference Engine Claims 4.5x Speedup Over MLX, 730 tok/s on MacBook — jimmykoppel · 2026-09-22
- Full-Parameter RL on TPUs: peano_ai Runs 310B MiMo-V2.6 Across 1,000+ TPUs — simonguozirui · 2026-09-22
- Cloudflare Python Workers go generally available after two-year preview — Simon Willison · 2026-09-22
- Fighting AI crawler traffic: beyond Turnstile, Cloudflare's AI Labyrinth as an option — fforres · 2026-09-22