← Browse

World's Fair 2024

Decoding the Decoder LLM without de code: Ishan Anand

Overview

This talk provides a deep dive into the inner workings of Large Language Models (LLMs) by dissecting GPT-2 small and reconstructing its functionality within a Microsoft Excel spreadsheet. The presentation aims to demystify LLMs for individuals without formal machine learning degrees, illustrating how text generation is fundamentally a complex mathematical problem. It covers the model's anatomy, its thought process through a virtual MRI, and concludes with a demonstration of AI "brain surgery" to alter its behavior.

Who should watch

Key takeaways

Notable quotes

*The core message I want to leave with is that to be a better AI engineer it does help to unlock the Black Box.*
*The more you can clear that up the more you can clear up misunderstandings.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.