| |
|
| |
Join the vLLM community & the NVIDIA Dynamo team for an evening of tech talks & conversation during vLLM conference week.
We'll dig into what it takes to serve LLMs efficiently at scale-covering the latest work across vLLM & NVIDIA Dynamo, from inference optimization & distributed serving to the practical challenges of running these systems in production. After the talks, stick around to connect with engineers, researchers, & builders from across the inference community over food & drinks.
Whether you're deep in the internals of an inference engine or just getting started with LLM serving, come learn, ask questions, & meet the people pushing this work forward.
When: Monday, August 24 6-9pm PT
Schedule:
6:00pm: Doors open
6:30-6:45pm: vLLM tech talk
6:45-7pm: Dynamo tech talk
7-9pm: Socialize & mingle!
Space is limited & registration is subject to approval-reserve your spot early.
|
|
|
|
|
|
|
|