Smokebench: Lightweight TUI for quick local LLM benchmarking
Ninja-5000 · reddit · 2026-08-21
Smokebench is an open-source, lightweight TUI for quick local LLM benchmarking (smoke-testing). It supports OpenAI/Anthropic compatible endpoints, covers 8 categories like math and code, offers LLM-as-judge scoring, TTFT/TPS metrics, and exportable logs. Works with Ollama, LM Studio, vLLM, etc.
More from coding & agent
- Single binary harness orchestrates Claude, GPT, DeepSeek, and GLM simultaneously — K_Kolomeitsev · 2026-08-21
- PuppyOne Uses Filesystem as Agent State, Bypassing Custom Stores — rohanpaul_ai · 2026-08-21
- Dev argues for upfront alignment over review in AI agent workflows — mattpocockuk · 2026-08-21
- Dev builds a game in 6 minutes using ChatGPT, showcasing AI speed — aziz4ai · 2026-08-21
- Using Agent Orchestration: Easy to use but slower and less deterministic — mattpocockuk · 2026-08-21
- LLM Financial Verification Benchmark: Deterministic 100%, Live LLM 29% — MuhammadMujtaba21 · 2026-08-21