← Browse

Europe 2026

Under 5 minutes to a deployed LLM endpoint — Audry Hsu, RunPod

Audry Hsu

Overview

This talk introduces RunPod as a cloud AI infrastructure platform designed to simplify GPU access and model deployment for developers. The core thesis is that managing complex infrastructure, especially GPU hardware, is a significant hurdle for builders, and RunPod aims to abstract this away. The platform allows users to bring their own code and models, whether private or open-source, and deploy them quickly, focusing on enabling developers to build applications rather than manage hardware.

Who should watch

Key takeaways

Notable quotes

*RunPod, we are a cloud AI infrastructure company. So, we have the hardware, we have the GPUs, and we make it easy for developers to deploy models.*
*We want to build we as software developers, we bring bring the value through the applications that we build, not from managing the infrastructure.*
*For a lot of teams, serverless is the fastest way if you want to start deploying a production-ready API.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.