vLLM and NVIDIA Co-host Meetup on Scaling LLM Inference Efficiency
vllm_project · x · 2026-08-12
The vLLM project is co-hosting a meetup with the NVIDIA Dynamo team in San Francisco on August 24.
The event will feature tech talks focused on serving LLMs efficiently at scale. Discussions will cover the latest work across vLLM and NVIDIA Dynamo, including inference optimization, distributed serving, and practical challenges of running these systems in production.
More from Companies & People
- Uber exits Serve Robotics stake, partner learns via public filing — HaktanSuren · 2026-08-12
- Palantir CEO: Off-the-Shelf LLMs Won't Work, Enterprises Need an Ontology Layer — eliano · 2026-08-12
- Will AI's New Security Threats Topple Cybersecurity Giants? One Investor Says No — prateekj · 2026-08-12
- Micro1 Ranks No. 37 on Inc. 5000, CEO Ali Ansari is Youngest in Top 50 — Exp_Mark · 2026-08-12
- Mistral Unveils European Sovereign AI Plan: In-Region Inference, Open Models, Long-Term Commitments — MistralAI · 2026-08-12
- Mistral Reaffirms Open-Source Platform Strategy: Choice and Flexibility for Enterprises — MistralAI · 2026-08-12