← Browse

Europe 2026

Lessons from Trillion Token Deployments at Fortune 500s — Alessandro Cappelli, Adaptive ML

Alessandro Cappelli , Adaptive ML

Overview

This talk argues that reinforcement learning (RL) is crucial for moving AI models, particularly large language models (LLMs), from pilot stages to production. The speaker posits that the common failure of GenAI pilots stems from the "myth of the last mile," where initial MVPs built on proprietary models or instruction fine-tuning lack systematic improvement pathways. RL, by its nature, allows for the mathematical integration of feedback, leading to more effective model steering and enabling scaled, cost-efficient, and faster deployments.

Who should watch

Key takeaways

Notable quotes

*95% of GenAI pilots fail to reach production.*
*RL is the one algorithm that will let you bring model into production in a systematic and industrialized way.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.