← Browse

World's Fair 2024

Multi model multimodal and multi agent innovations in Azure AI: Cedric Vidal

Overview

This talk showcases advancements in Azure AI, focusing on multimodal and multi-agent capabilities. It highlights how new models and tools, such as GPT-4o and Phi-3 Vision, can process and reason across text, vision, and speech. The presentation emphasizes practical applications and the integration of these technologies within Azure AI Studio for building sophisticated AI solutions.

Who should watch

Key takeaways

Notable quotes

*The model understands natively both pixels and text and in its internal representation has the same vectors for the same concepts.*
*This will make the world more inclusive.*
*I didn't code a single line.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.