

AI Scale Talks EP.4 | You Built the Token Factory. Who Runs It?
AI Scale Talks is a four-part live webinar series from Lablup. The series theme is "From Cell to Factory": each week we scale up one level, from a single inference engine to full-scale AI infrastructure operations.
EP.4 starts with Continuum Router and Continuum Hub. Most teams now run several model providers at once, and the token has become the unit everyone measures by. It is also the thing nobody owns: usage is invisible, cost leaks, and policy sits wherever the call happened to be made. This session covers the two layers that close that gap: Continuum Router on the request path, and Continuum Hub, which meters, governs, and bills a fleet of routers from the center.
➡️ Format
40-minute session + 10-minute live Q&A
Live on Zoom. The join link is shared with registered guests.
📆 All episodes air Wednesdays at 9:00 AM PT.
EP.1 | Aug 26 | High-performance LLM/VLM inference with MLxcel engine, Jeongkyu Shin
EP.2 | Sep 2 | Accelerating LLM Inference: From Speculative Decoding to Diffusion LLMs, Junbum Lee
EP.3 | Sep 9 | Deploying an AI Platform on Infrastructure You Don’t Control, Jonghyun Park