

vLLM Meetup Toronto
Join us for the vLLM x Cohere meetup in Toronto, an evening for AI engineers, researchers, and infrastructure builders.
This is a chance to connect with vLLM maintainers and the Cohere team building and serving models for production-scale workloads. We'll cover how the vLLM ecosystem and Cohere's open-weights work and upstream contributions fit together.
The session goes deep on inference: the vLLM roadmap, serving for agentic workloads, and advanced KV cache offloading. Roger Wang, cofounder of Inferact and a core vLLM maintainer, will give technical updates, and we'll close with a live panel on where inference is heading and what's next for open source AI. Additional speakers from vLLM and Cohere will be announced soon.
We'll finish the evening with networking, drinks, and bites.
Schedule
5:00pm - Doors open
5:30pm - 7:00pm - Talks & Panel Q&A
7:00pm - 9:00pm - Networking Reception