Build Evals That Actually Matter - Nick Ung & Akshay Sharma, Lyft
This talk focuses on the critical importance of building effective evaluation systems for AI models, moving beyond superficial metrics to create benchmarks that genuinely reflect real-world performance and user needs. The presenters emphasize that robust evals are essential for shipping reliable AI products and driving meaningful improvements in model quality.
World's Fair 2026 38 min