← Browse

World's Fair 2025

OpenAI on Securing Code-Executing AI Agents — Fouad Matin (Codex, Agent Robustness)

Agent Robustness)

Overview

This talk addresses the critical security and safety considerations for AI agents capable of executing code. As AI models become increasingly proficient at writing and running code, the focus shifts from mere capability to responsible deployment and robust guardrails. The presentation emphasizes that code execution is becoming a standard feature for AI agents, moving beyond traditional software engineering tasks to achieve objectives more efficiently across various applications.

Who should watch

Key takeaways

Notable quotes

*It's not just actually about writing code but it's about achieving the objective most efficiently.*
*The new constraint isn't just can these models do things but actually what should they be able to do and what should the guardrails be.*
*The most common one, something we think about consistently is prompt injection and data exfiltration.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.