Featured in
Bond AI - San Francisco and Bay Area
Introduction to AI Evals | San Francisco
Registration
Past Event
About Event
Join our interactive AI Evaluations Workshop to learn practical skills in evaluating large language models (LLMs) and AI agents!
We’ll cover key concepts including using LLM APIs, understanding jailbreaks, why evaluations (evals) are essential yet challenging, and insights from Anthropics “Alignment Faking Paper”. The session features live coding exercises based on the ARENA materials (section 3.1), using Python and Jupyter notebooks.
Prior Python experience is required. Ideal for AI practitioners, researchers, or enthusiasts looking to deepen their understanding of AI safety and evaluations.