← Browse

World's Fair 2026

Build Evals That Actually Matter - Nick Ung & Akshay Sharma, Lyft

Nick Ung , Akshay Sharma , Lyft

Overview

This talk focuses on the critical importance of building effective evaluation systems for AI models, moving beyond superficial metrics to create benchmarks that genuinely reflect real-world performance and user needs. The presenters emphasize that robust evals are essential for shipping reliable AI products and driving meaningful improvements in model quality.

Who should watch

Key takeaways

Notable quotes

The presenters stressed the need for evals that *actually matter* in reflecting real-world utility.
Building good evals is presented as a core part of the product development lifecycle for AI.

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.