More MLOps.community episodes

Performance Optimization and Software/Hardware Co-design across PyTorch, CUDA, and NVIDIA GPUs thumbnail

Performance Optimization and Software/Hardware Co-design across PyTorch, CUDA, and NVIDIA GPUs

Published 24 Feb 2026

Duration: 01:25:49

The podcast dives into AI development, software engineering, and GPU innovation, focusing on efficient workloads and the trade-offs between quick solutions and scalable systems.

Episode Description

March 3rd, Computer History Museum CODING AGENTS CONFERENCE, come join us while there are still tickets left.https://luma.com/codingagentsChris Fregly...

Overview

The podcast covers multiple topics related to product innovation, software development, and engineering practices, focusing on trends and challenges in AI and machine learning. It examines the use of SageMaker HyperPods, which employ pre-warmed GPUs to improve efficiency for AI workloads, and discusses the rise of "throwaway" applications designed for specific, short-term needs without long-term maintenance. The conversation also addresses the trade-off between quick, functional solutions and scalable, robust systems, emphasizing the role of software engineers in developing reliable production-grade applications. Additionally, the episode explores the use of AI tools in code generation and debugging, including a feature called the "playground skill" for visualizing code flow.

The discussion extends to issues with GPU hardware and limitations in AI infrastructure, highlighting the need for better documentation and transparency. It also touches on the growing interest in optimizing AI models for specific applications and newer hardware such as NVIDIA's Blackwell. The author reflects on writing a book focused on co-design principles that integrate hardware, software, and algorithms, and underscores the importance of open-source tools and community collaboration in advancing AI development and deployment.

Recent Episodes of MLOps.community

14 Sept 2026 Why Cost Per Million Tokens Is A Useless KPI?

"AI's exponential growth in cloud services demands new FinOps strategies to manage unbounded costs, unpredictable usage, and real-time tracking, requiring adapted SRE/DevOps principles and use-case-specific financial modeling."

27 Jul 2026 What an Anthropic Engineer Thinks About MCP

"SDKs now see hundreds of millions of downloads annually, with a focus on minimal, extensible designs and a major MCP update shifting to stateless protocols for scalability, balancing simplicity with complexity while prioritizing stability and future-proofing."

More MLOps.community episodes