One Year of LLM Inference Metadata Released: 6.12 Billion Requests for Serving Research
generativist · x · 2026-09-21
Research group 1a1a11a announced the release of one year of LLM inference metadata traces covering 6.12 billion requests, built by grad student William Nixon in collaboration with Jon Durbin and Chutes AI. The dataset targets real-world LLM serving workload understanding, system design, and infrastructure optimization.
Related event: Harvard Releases Dataset of 6.12 Billion Real-World LLM Inference Requests(3 posts)→
More from Infra
- Grass network audited: 3M+ users, $32.1M revenue serving AI training data — Ronangmi · 2026-09-22
- SemiAnalysis: mapping MoE models onto inference hardware — zephyr_z9 · 2026-09-22
- Huawei's HiZQ is real HBM; DeepSeek's designs all target bandwidth, not storage — zephyr_z9 · 2026-09-22
- Halo RL training platform launches with SGLang as primary rollout engine — ying11231 · 2026-09-22
- SiliconBench: speed, memory and fidelity of nine LLM engines on unified-memory desktops — PennState · 2026-09-22
- Engram retrieval won't replace FFNs, but cutting 40-50% of HBM needs is the real win — bookwormengr · 2026-09-22