Cover Image for Reward Hacking Defense β€” Breakfast
Cover Image for Reward Hacking Defense β€” Breakfast
Avatar for META Γ— Hugo Β· NYC

Reward Hacking Defense β€” Breakfast

Registration
Approval Required
Your registration is subject to host approval.
Welcome! To join the event, please register below.
About Event

β€‹πŸ“ About the Event

​

​Breakfast on what the reward model rewards but you did not intend. GRPO, DPO, and what breaks underneath them: the failures automated test suites do not catch because the model learned to satisfy the metric rather than the task.

​

β€‹βœ¨ What to Expect

​

​‒ Breakfast at Madman Espresso, an early start and a short session

​‒ Reward hacking, GRPO and DPO, and the failures that survive automated QA

​‒ Sandboxed execution and dual-engine consensus, in practice not theory

​‒ Hosted by Evan Ward, Hugo

​

β€‹πŸ§  About Hugo Inc.

​

​Hugo Inc builds managed, expert-density teams for the labs training and evaluating frontier models, and has been a META partner since 2018. Not a crowd platform and not anonymous click-workers: 730+ vetted specialists, 100 percent university-educated, 64 percent holding four-year STEM or CS degrees, working inside your sandbox rather than around it.

​

​The work: reasoning-trace auditing, process rewards and RLHF preference data, expert SFT authoring, multimodal grounding, agentic and coding evaluation, and adversarial red teaming. 98.90 percent average accuracy, 99 percent+ first-pass quality, and sub-1.5 percent monthly attrition at a 3.5-year average tenure, so the same reviewers stay on your pipeline.

​

​Full rundown of what we do for labs like yours: hugoinc.com/industry/frontier-reasoning-labs

Location
Madman Espresso 11 Ave
311 11th Ave, New York, NY 10001, USA
Avatar for META Γ— Hugo Β· NYC