← Browse

Europe 2026

$1 AI Guardrails: The Unreasonable Effectiveness of Finetuned ModernBERTs – Diego Carpentero

Diego Carpentero

Overview

This talk addresses the escalating sophistication of AI attacks, moving beyond simple prompt injection to complex exploits targeting LLM interfaces, data, and even internal mechanisms. It proposes a practical, low-latency, self-hosted defensive layer built by fine-tuning a modern encoder model, specifically ModernBERT, to act as a guardrail against these threats. The approach emphasizes efficiency and cost-effectiveness, aiming for a defense under a dollar per instance.

Who should watch

Key takeaways

Notable quotes

*The LLMs, they have no native separation of concerns between the system controls and the data.*
*We're not building defensive layers to pass a security audit. We have to build safety mechanisms that protect machines, human and humans and society.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.