Cover Image for Learn Generative AI: Evaluation and Monitoring for LLM Systems
Cover Image for Learn Generative AI: Evaluation and Monitoring for LLM Systems
Avatar for Open Source for AI
Presented by
Open Source for AI
Providing all developers the resources to understand, use, and contribute to the development and direction of AI
Hosted By

Learn Generative AI: Evaluation and Monitoring for LLM Systems

YouTube
Registration
Past Event
Welcome! To join the event, please register below.
About Event

Sponsored by Cracking Gen AI.

This week's speaker will be Oscar Courbit, Member of Technical Staff at Clarity Care AI. Oscar has previously been a founding ML Engineer at AI Researcher at MIT.

Talk Abstract

LLMs started as tools we called. Now they’re agents that decide: for our users, for our engineering, for the industry. As LLM systems gain autonomy, two concerns that used to be separate, evaluation and monitoring, converge into a single continuous loop.

Drawing from production experience building AI-powered healthcare systems, this talk presents a practical framework for evaluating and monitoring LLM systems at every stage of autonomy. We’ll cover how to define what failure actually means for your users (it’s not what you think), how to build a root cause taxonomy that tells you exactly where to invest, and how to turn a manual investigation into a self-improving monitoring pipeline. We’ll also explore what changes when the AI becomes more autonomous: agents evaluating agents, step-level quality signals, and the emerging pattern of self-healing systems.

Anticipated Agenda

6:00 - 6:05: Introduction

6:05 - 6:45: Presentation

6:45 - 7:00: Q/A

Avatar for Open Source for AI
Presented by
Open Source for AI
Providing all developers the resources to understand, use, and contribute to the development and direction of AI
Hosted By