← Browse

Code 2025

Context Platform Engineering to Reduce Token Anxiety — Val Bercovici, WEKA

Val Bercovici

Overview

This talk introduces context platform engineering as a method to reduce token anxiety and improve AI agent performance. The core thesis is that by optimizing how context is managed and cached, developers can significantly increase KV cache hit rates, leading to more efficient and cost-effective AI systems. The presenters announce the open-sourcing of their context platform engineering toolkit, designed to help engineers achieve these optimizations.

Who should watch

Key takeaways

Notable quotes

*KV cache hit rates are the single most important metrics for production grade AI agents.*
*Context platform engineering quite simply maximizes KV cache hit rates in a very straightforward manner.*
*We're thrilled to be announcing the open sourcing of this context platform engineering toolkit today.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.