← Browse

World's Fair 2025

RL for Autonomous Coding — Aakanksha Chowdhery, Reflection.ai

Aakanksha Chowdhery

Overview

This talk explores the progression of large language models (LLMs) from early scaling laws to the current frontier of autonomous coding. It highlights how techniques like chain-of-thought prompting and reinforcement learning with human feedback have improved LLM capabilities. The core thesis is that reinforcement learning, particularly in domains with automated verification like coding, represents the next significant step in scaling LLM performance and building more intelligent systems.

Who should watch

Key takeaways

Notable quotes

*The next era from this year is really the era of experience which was which will lead us to super intelligence.*
*Reinforcement learning will be a fundamental component in building super intelligent systems especially in areas where we have automated verification.*
*Coding is one of those domains where we do have the capability to verify and that gives us tremendous advantage in terms of building super intelligence on top of autonomous coding.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.