← Browse

Session brief

Fixing bugs in Gemma, Llama, & Phi 3: Daniel Han

18 min

Overview

This talk addresses common bugs and issues encountered when fine-tuning and deploying open-source large language models, specifically focusing on Gemma, Llama 3, and Phi 3. It provides practical solutions and best practices to ensure successful model training and inference, highlighting the importance of careful attention to tokenization, model templates, and export formats.

Who should watch

Key takeaways

Notable quotes

*Please check before you fine tune if you're using double BOS tokens.*
*The Llama 3 chat template will not work for the base model.*
*The pad token and the EOS token must not be the same.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.