vLLM Conference Agenda Released: NVIDIA, AMD, Google to Discuss the Future of AI Inference
vllm_project · x · 2026-08-05
vLLM has officially released the full agenda for its first conference, taking place August 24–26 in San Francisco at the Ray Summit. The speaker lineup features key engineers and researchers from Inferact, NVIDIA, AMD, Google TPU, Meta, and more.
Key topics include:
- vLLM Roadmap: Deep dives into the Flat Model and Model Runner V2 migrations, alongside the future of the engine and ecosystem.
- Hardware & Optimization: Covering accelerator support, training integrations, and production-scale inference work.
Related event: vLLM Announces Inaugural AI Inference Summit in San Francisco for August(2 posts)→
More from Infra
- Garage Solar-Powered 384GB Xeon Rig for Remote AI Coding — angadsg · 2026-08-05
- AMD CEO Lisa Su to Analyst: Your Data Center AI Number is Probably Too Low — firstadopter · 2026-08-05
- Open-Sourced Recipe: Running a 27B Local Agent 24/7 on a Single RTX 5090 — max_paperclips · 2026-08-05
- Europe Pledges €30B for AI Gigafactories, Only €1B Actually Committed — sanjaykalra · 2026-08-05
- Modal Optimizes Serverless Architecture to Reduce Network Latency — AAAzzam · 2026-08-05
- Elon Musk: AI Memory Demand Growing Over 200% Annually, Supply Lagging — firstadopter · 2026-08-05