▶ Watch ↗AI Engineer Europe 202643:53
Bio, Work & Ideas
Diego Carpintero is an AI engineer developing self-hosted AI guardrails that make defenses for language-model applications fast, affordable, and practical to operate independently. His security work addresses a fundamental architectural vulnerability: language models process trusted developer instructions and untrusted external information together, allowing ordinary-looking text to redirect automated decisions and actions.
At AI Engineer Europe 2026, Carpintero mapped how prompt injection expands beyond chatbot inputs into poisoned retrieval documents, deceptive Model Context Protocol tool descriptions, adversarial token sequences, compromised dependencies, and autonomous agents. He emphasizes that human approval can fail when users see an innocuous tool summary while the model receives hidden malicious instructions.
His ModernBERT guardrails demonstration connects inexpensive local deployment with a larger responsibility: protecting sensitive information, preventing unauthorized actions, and limiting the human consequences of manipulated automated decisions.