First vLLM Conference at Ray Summit

Hosted byInferact logo

August 24 - 26, San Francisco

The first vLLM Conference at Ray Summit brings together the engineers and researchers advancing open, high-performance inference.

Register here

Interested in sponsoring or speaking? Email events@vllm.ai.

Speaker lineup

Builders behind the vLLM ecosystem

Sessions cover the vLLM roadmap, hardware backends, training and serving integrations, and production-scale inference work.

Woosuk Kwon profile photo
Inferact logo

Woosuk Kwon

Cofounder / CTO, Inferact; Creator / Core Maintainer, vLLM

Siyuan Fu profile photo
NVIDIA logo

Siyuan Fu

Senior Engineer, NVIDIA

Douglas Lehr profile photo
AMD logo

Douglas Lehr

Principal Engineer, AMD

Qi Zhou profile photo
Google TPU logo

Qi Zhou

Senior Staff Software Engineer, Google TPU

Michael Goin profile photo
Red Hat logo

Michael Goin

Senior Principal Machine Learning Engineer, Red Hat

Richard Zou profile photo
PyTorch logo

Richard Zou

Senior Staff Software Engineer, Meta

Yifan Qiao profile photo
Inferact logo

Yifan Qiao

Member of Technical Staff, Inferact

Sami Jaghouar profile photo
Prime Intellect logo

Sami Jaghouar

Research Engineer, Prime Intellect

Harry Mellor profile photo
Hugging Face logo

Harry Mellor

Machine Learning Engineer, Hugging Face

Tun Jian Tan profile photo
Embedded LM logo

Tun Jian Tan

Senior Tech Lead - AI/LLM System, Embedded LM

Zachary Xi profile photo
Inferact logo

Zachary Xi

Product, Inferact

Greg Pereira profile photo
Red Hat logo

Greg Pereira

Sr. Machine Learning Engineer, Red Hat

Zijing Liu profile photo
Inferact logo

Zijing Liu

Member of Technical Staff, Inferact

Jeffrey Wang profile photo
Anyscale logo

Jeffrey Wang

Software Engineer, Anyscale

Sara Smoot profile photo
Google DeepMind logo

Sara Smoot

Software Engineering Manager, Google DeepMind / Gemma

Fred Wang profile photo
Together AI logo

Fred Wang

ML Research Engineer, Together AI / LightSeek

Chendi Xue profile photo
Intel logo

Chendi Xue

Machine Learning Engineer, Intel

Debarshi Raha profile photo
DigitalOcean logo

Debarshi Raha

VP / Fellow Engineer, DigitalOcean

Conference schedule

Two days of vLLM sessions

Explore talks spanning the vLLM roadmap, hardware backends, agentic serving, training, and production inference.

Day 1

Tuesday, August 25

8:30 AMRegistration + Networking
9:30 AMKeynotes
11:30 AMNetworking Lunch
12:30 PMState of vLLM 2026Woosuk Kwon and Zachary Xi · Inferact
1:00 PMBreak
1:15 PMNVIDIA and vLLM Full-Stack Collaboration for DeepSeek and MiniMax PerformanceSiyuan Fu · NVIDIA
1:45 PMBreak
2:00 PMPyTorch <3 vLLMRichard Zou · PyTorch / Meta
2:30 PMBreak
2:45 PMAccelerating Large Language Models: vLLM on Google TPUsQi Zhou · Google TPU and Sara Smoot · Gemma
3:15 PMBreak
3:30 PMHow a Transformers Model Loads in vLLMHarry Mellor · Hugging Face
4:00 PMBreak
4:15 PMProduction-Grade Distributed Inference with llm-dGreg Pereira and Michael Goin · Red Hat
4:45 PMBreak
5:00 PMGLM-5 RL Training with vLLMSami Jaghouar · Prime Intellect
5:30 PMRay After Party

Day 2

Wednesday, August 26

8:30 AMRegistration + Breakfast + Networking
9:30 AMKeynotes
11:30 AMNetworking Lunch
12:30 PMServing vLLM on Agentic Production WorkloadsYifan Qiao and Zijing Liu · Inferact
1:00 PMBreak
1:15 PMBeyond Porting: High-Performance vLLM Inference on AMD ROCmDouglas Lehr · AMD
1:45 PMBreak
2:00 PMBeyond Tokens: Making vLLM Stateful and Agent-Ready with Agentic APITun Jian Tan · Embedded LM
2:30 PMBreak
2:45 PMKV Cache Aware Serving for Agentic and RL Workloads with Ray Serve LLMJeffrey Wang · Anyscale
3:15 PMBreak
3:30 PMServing vLLM on Intel GPUs, CPUs, and GaudiChendi Xue · Intel
4:00 PMBreak
4:15 PMTorchSpec: Speculative Decoding Training at ScaleFred Wang · Together AI / LightSeek
4:45 PMBreak
5:00 PMHow Open-Source vLLM Topped the Artificial Analysis LeaderboardDebarshi Raha · DigitalOcean
5:30 PMEvent Ends

Ecosystem represented

Companies building around vLLM

Inferact logoNVIDIA logoAMD logoGoogle TPU logoIntel logoPyTorch logoMeta logoDigitalOcean logoRed Hat logoHugging Face logoPrime Intellect logoAnyscale logoGoogle DeepMind logoGoogle Gemma logoTogether AI logoLightSeek logoEmbedded LM logo