← Browse

World's Fair 2025

RFT, DPO, SFT: Fine-tuning with OpenAI — Ilan Bigio, OpenAI

Ilan Bigio

Overview

This talk explores various fine-tuning techniques for OpenAI models, including Supervised Fine-Tuning (SFT), Direct Preference Optimization (DPO), and Reinforcement Fine-Tuning (RFT). The presenter, Ilan Bigio from OpenAI's developer experience team, emphasizes that fine-tuning is a specialized tool for optimizing models beyond what prompt engineering can achieve, particularly for specific domains or behaviors. The discussion covers the data requirements, use cases, and limitations of each method, offering practical examples and best practices.

Who should watch

Key takeaways

Notable quotes

*Fine-tuning is really just continued training that optimizes a model for a given domain.*
*Prompting is kind of like a set of tools like a hammer, um, pliers, whatever.*
*Fine-tuning is like a CNC machine, right? you have uh like it's a way higher upfront investment.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.