
vLLM Inference Meetup Bengaluru
Description
Join us for a deep dive into the engine room of vLLM and llm-d AI inferencing, where we will focus on the architecture, optimizations, and raw engineering required to run inference at scale. Whether you’re looking to squeeze every last token out of your GPU cluster or you're curious about the latest commits to the vLLM and llm-d ecosystems, this is the room you want to be in. What to Expect Who Should Attend Agenda (Subject to More Awesomeness) Important information Registration closes 24 hours before the event. We cannot admit unregistered attendees. Please bring a photo ID to verify your registration on arrival.
Event location
Let your network know you`re going
Share this event to start conversations, invite colleagues, and connect before it begins.
Artificial Intelligence Platform
Free
