

Workshop: Build your eval foundations
Every great agent improves through a continuous loop: observe behavior, evaluate performance, and iterate with confidence.
In this hands-on workshop, you'll build that loop from scratch.
Working step by step alongside the Braintrust team, you'll instrument a real agent, inspect traces, create your first evals and scorers, and learn how to turn agent behaviour into measurable improvements. This session gives you the building blocks you'll use every time you build and ship an agent.
You'll leave with a practical workflow you can use to:
Observe what your agent is doing
Build evals that measure quality
Continuously improve agent performance
Agenda
3:00 PM Registration
3:30 PM Intro by Jess Wang
4:00 PM Hands-on workshop led by Curtis Galione
6:00 PM Rooftop drinks and canapés
Bring your laptop and stay for the rooftop drinks.