Context Platform Engineering to Reduce Token Anxiety — Val Bercovici, WEKA
This talk introduces context platform engineering as a method to reduce token anxiety and improve AI agent performance. The core thesis is that by optimizing how context is managed and cached, developers can significantly increase KV cache hit rates, leading to more efficient and cost-effective AI systems. The presenters announce the open-sourcing of their context platform engineering toolkit, designed to help engineers achieve these optimizations.
Code 2025 24 min