Cover Image for July 24: Embodied Reasoning with World Models by Yilun Du
Cover Image for July 24: Embodied Reasoning with World Models by Yilun Du
Avatar for Video Model Journal Club

July 24: Embodied Reasoning with World Models by Yilun Du

Register to See Address
Registration
Past Event
Please click on the button below to join the waitlist. You will be notified if additional spots become available.
About Event

​Abstract: I'll present a couple methods showing how we can effectively reason with video models. I'll first talk about how we can directly reason with video models. I'll then talk about how video models can be combined with VLMs for hybrid reasoning and with policies for low level reasoning.

​Speaker: Yilun Du — Assistant Professor at Harvard University. Previously at MIT CSAIL, where his research focuses on generative models, world models, and embodied AI for robot learning and planning.

​Every week we pick one paper and go deep — video generation, world models, physical reasoning, diffusion, flow matching, and everything in between.

​Website: https://journal.video-reason.com/To join over zoom, please subscribe to get zoom link: Open Google Form

​Co-Hosted by 1943

​1943 is an invite-only community for researchers and builders who value intellectual depth. We host closed-door discussions and research talks in the Bay for those pushing the frontier. Follow us on X or Linkedin for future research talks and roundtable discussions!

​Co-Hosted by Moonlake AI

​Moonlake is a world model a simulation infrastructure platform to accelerate physical AI deployment. Follow them on X or Linkedin__.

​Co-Hosted by Tripo AI

​Tripo is an AI 3D generation platform that turns text and images into production-ready 3D models. Follow them on X or LinkedIn for the latest in generative 3D!

Location
Please register to see the exact location of this event.
Avatar for Video Model Journal Club