

CAIA Speaker Event: Jerry Wei (Anthropic)
Here are some details on Caltech AI Alignment's next speaker event:
Who: Jerry Wei (in person), Anthropic
When: May 29th at 4 pm PT
Where: Broad 100
What: Jerry Wei is an AI researcher at Anthropic (formerly Google DeepMind). His talk will cover how a frontier lab decides a model is safe to ship. Using Anthropic's February 2026 risk report on biology, he'll walk through what the Responsible Scaling Policy commits them to, how capability evals are designed, how defenses like Constitutional Classifiers are built and stress-tested, and why safety training alone can't be counted on to degrade dangerous capabilities away.
No specific technical background is required - we welcome all interested students who are eager to learn! As with all CAIA events, we will have pizza and boba!