← Browse

World's Fair 2026

Your Agent Is Wasting Tokens and You Don't Know It - Erik Hanchett, AWS

Erik Hanchett

Overview

This talk addresses the significant token costs associated with using AI agents, particularly large language models, and provides practical strategies for reducing these expenses. The core thesis is that by implementing specific techniques, developers can optimize agent performance and cost-efficiency without sacrificing functionality.

Who should watch

Key takeaways

Notable quotes

*You want to use multiple different models based on the use case.*
*If you can find any way that where you have this tool result that you don't necessarily send it on every single call back to the large language model, that will save a lot of tokens for you.*
*So always set a max iterations of how many times it will loop.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.