← Browse

Europe 2026

Accelerating AI on Edge — Chintan Parikh and Weiyi Wang, Google DeepMind

Chintan Parikh , Weiyi Wang

Overview

This talk focuses on accelerating AI deployment on edge devices, highlighting Google DeepMind's Gemma models and the Lite RT framework. The core thesis is that by optimizing models for edge hardware and leveraging a unified cross-platform architecture, developers can achieve significant performance gains, enhanced privacy, and reduced latency for a wide range of AI applications.

Who should watch

Key takeaways

Notable quotes

*Running on edge has many benefits. Certainly like latency is important for those who are keen.*
*The big evolution with Gemini 4 is going to be really moving from like chatbot type capabilities to more autonomous agents that also support reasoning capabilities.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.