← Browse

Europe 2026

Your coding agent doesn't always follow your rules — Talha Sheikh, Checkout.com

Talha Sheikh

Overview

This talk addresses the common issue where AI coding agents, despite appearing to complete tasks, often produce outputs that fail upon execution or do not meet specifications. The core argument is that the value is shifting from the agent's ability to generate code to the developer's ability to design and implement robust verification systems, or harnesses, that ensure the agent's output is reliable and deterministic.

Who should watch

Key takeaways

Notable quotes

*The agent says it's done but you have to check it anyway because there's nothing else that can check it for you.*
*It's not about whether Claude can actually do the task, it's about trust.*
*The value is not in the code that we create, but it's actually now in reality is what we're seeing here is the verification that we design.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.