More Practical AI episodes

Controlling AI Models from the Inside thumbnail

Controlling AI Models from the Inside

Published 20 Jan 2026

Duration: 2635

The podcast delves into the AI safety crisis, discussing ongoing struggles with AI-generated harm, the limitations of current security measures, and emerging solutions for real-time monitoring and more sophisticated safety protocols.

Episode Description

As generative AI moves into production, traditional guardrails and input/output filters can prove too slow, too expensive, and/or too limited. In this...

Overview

The podcast explores the ongoing challenges in AI safety, particularly the risks of AI systems generating harmful or unintended content such as violence, pornography, or dangerous advice. It differentiates between using AI for security and ensuring AI systems themselves are secure, stressing the importance of proactive safety measures beyond basic input and output filtering. Current approaches are criticized for being reactive, often relying on post-hoc analysis of outputs and struggling with detecting harmful content in complex media like video and audio.

The discussion highlights emerging solutions that use internal model instrumentation to identify unsafe behavior in real-time, offering a more efficient and scalable alternative. It also addresses the value of interpretability in AI, the need for layered defense strategies, and the potential of edge devices to support safety mechanisms with lower computational requirements. The conversation touches on economic and practical barriers to implementing strong safety measures and the difficulty of tailoring these systems to industry-specific needs, while envisioning a future of more adaptable and context-aware AI security frameworks.

Recent Episodes of Practical AI

17 Sept 2026 How to get discovered in AI search

"Explores AI-driven search's impact on digital marketing, shifting from SEO to strategies like AEO and GEO, and challenges in adapting to AI's unique retrieval methods, brand visibility, and evolving content needs."

10 Sept 2026 Computer-Use Agents and the Future of the Agentic Internet

"AI's real-world impact is explored, covering its role in daily life, work, and creativity, with a focus on evolving AI agents, trust challenges, B2C vs. B2B applications, future agentic interfaces, and ethical concerns like centralized control and rapid advancements."

28 Aug 2026 Building the Foundation for the Agentic AI Era

"AI leader Angie Jones detailed her work at IBM, Twitter, and Block - including training 12,000 employees on AI agents like Goose - while advocating for deeper AI integration, open standards (e.g., MCP), and ethical, human-augmenting AI development."

25 Aug 2026 AI Proficiency: From Users to Builders

"AI's real-world impact in business and daily life, focusing on strategic adoption, workforce transformation, and enhancing human roles through adaptability and measurable value."

More Practical AI episodes