Cover Image for Demystifying Video Reasoning by Ruisi Wang
Cover Image for Demystifying Video Reasoning by Ruisi Wang
Avatar for Video Model Journal Club
Hosted By
54 Went

Demystifying Video Reasoning by Ruisi Wang

Zoom
Registration
Past Event
Welcome! To join the event, please register below.
About Event

​Abstract: Recent advances in video generation have revealed an unexpected phenomenon: diffusion-based video models exhibit non-trivial reasoning capabilities. We challenge the Chain-of-Frames assumption and uncover a fundamentally different mechanism — Chain-of-Steps (CoS), where reasoning emerges along the diffusion denoising steps. We identify several emergent reasoning behaviors: working memory, self-correction, and perception before action.

​Speaker: Ruisi Wang — Researcher with a background in computer science from Nanyang Technological University, working on computer vision, video reasoning, and spatial intelligence.

​All sessions: https://luma.com/video-model Register on this page to receive the Zoom join link (sent automatically with your confirmation).

​Subscribe to our mailing list for weekly talk announcements: Open Google Form

Avatar for Video Model Journal Club
Hosted By
54 Went