MLPerf Inference v6.1 draws record 30 submitters, adds agentic inference benchmarks

TheKanter · x · 2026-09-17

MLCommons released MLPerf Inference v6.1 results with a record 30 submitting organizations and up to 5.7X performance gains over last year. The release adds two new tests reflecting agentic AI deployment trends, including an End-to-End RAG benchmark covering the full pipeline of embedding, vector retrieval, re-ranking, and LLM reasoning, plus first peer-reviewed results for several new AI platforms.

Related event: MLPerf Inference v6.1 Adds First Agentic AI Benchmarks(2 posts)→

Original post →

More from Infra

Infra channel →