← Browse

World's Fair 2025

Latent Space Paper Club: AIEWF Special Edition (Test of Time, DeepSeek R1/V3) — VIbhu Sapra

VIbhu Sapra

Overview

This talk reviews recent advancements in AI research, focusing on the DeepSeek models and the evolution of training methodologies. It introduces a new "Test of Time" paper club initiative aimed at systematically covering foundational AI concepts. The discussion highlights how reinforcement learning and extended inference time are enabling models to develop advanced reasoning capabilities, leading to significant performance improvements.

Who should watch

Key takeaways

Notable quotes

*The Deepseek team says our goal is to explore the potential of LLMs to develop reasoning capabilities without any supervised data focusing on their self-evolution through a pure RL process.*
*This moment is not only an aha moment for the model but also for the researchers observing its behavior. It underscores the power and beauty of reinforcement learning.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.