Cover Image for vLLM Meetup Toronto
Cover Image for vLLM Meetup Toronto
Avatar for vLLM Meetups and Events
Join the vLLM community to discuss optimizing LLM inference!

vLLM Meetup Toronto

Register to See Address
Toronto, Canada
Registration
Approval Required
Your registration is subject to host approval.
Welcome! To join the event, please register below.
About Event

Join us for the vLLM x Cohere meetup in Toronto, an evening for AI engineers, researchers, and infrastructure builders.

This is a chance to connect with vLLM maintainers and the Cohere team building and serving models for production-scale workloads. We'll cover how the vLLM ecosystem and Cohere's open-weights work and upstream contributions fit together.

The session goes deep on inference: the vLLM roadmap, serving for agentic workloads, and advanced KV cache offloading. Roger Wang, cofounder of Inferact and a core vLLM maintainer, will give technical updates, and we'll close with a live panel on where inference is heading and what's next for open source AI. Additional speakers from vLLM and Cohere will be announced soon.

We'll finish the evening with networking, drinks, and bites.

Schedule

  • 5:00pm - Doors open

  • 5:30pm - 7:00pm - Talks & Panel Q&A

  • 7:00pm - 9:00pm - Networking Reception

Location
Please register to see the exact location of this event.
Toronto, Canada
Avatar for vLLM Meetups and Events
Join the vLLM community to discuss optimizing LLM inference!