More MLOps.community episodes

Performance Optimization and Software/Hardware Co-design across PyTorch, CUDA, and NVIDIA GPUs thumbnail

Performance Optimization and Software/Hardware Co-design across PyTorch, CUDA, and NVIDIA GPUs

Published 24 Feb 2026

Duration: 01:25:49

The podcast dives into AI development, software engineering, and GPU innovation, focusing on efficient workloads and the trade-offs between quick solutions and scalable systems.

Episode Description

March 3rd, Computer History Museum CODING AGENTS CONFERENCE, come join us while there are still tickets left.https://luma.com/codingagentsChris Fregly...

Overview

The podcast covers multiple topics related to product innovation, software development, and engineering practices, focusing on trends and challenges in AI and machine learning. It examines the use of SageMaker HyperPods, which employ pre-warmed GPUs to improve efficiency for AI workloads, and discusses the rise of "throwaway" applications designed for specific, short-term needs without long-term maintenance. The conversation also addresses the trade-off between quick, functional solutions and scalable, robust systems, emphasizing the role of software engineers in developing reliable production-grade applications. Additionally, the episode explores the use of AI tools in code generation and debugging, including a feature called the "playground skill" for visualizing code flow.

The discussion extends to issues with GPU hardware and limitations in AI infrastructure, highlighting the need for better documentation and transparency. It also touches on the growing interest in optimizing AI models for specific applications and newer hardware such as NVIDIA's Blackwell. The author reflects on writing a book focused on co-design principles that integrate hardware, software, and algorithms, and underscores the importance of open-source tools and community collaboration in advancing AI development and deployment.

Recent Episodes of MLOps.community

27 Jul 2026 What an Anthropic Engineer Thinks About MCP

"SDKs now see hundreds of millions of downloads annually, with a focus on minimal, extensible designs and a major MCP update shifting to stateless protocols for scalability, balancing simplicity with complexity while prioritizing stability and future-proofing."

20 Jul 2026 The Creator of FastMCP Explains the Future of MCP

"Fast MCP streamlined the Multi-Chat Protocol, dominating the market with simplicity and efficiency, while evolving to support interactive UI apps, Python-based token-efficient interfaces, and addressing security and scalability challenges, with AI tools enhancing personal and professional workflows."

13 Jul 2026 What Happens When Every Developer Has 20 AI Agents?

"Modern software development faces bottlenecks from limited human resources and AI-driven shifts, transforming productivity, SaaS models, and workflows while straining infrastructure and open-source ecosystems."

6 Jul 2026 AI Agents Should Be Treated Like Hackers

Integrating AI agents with enterprise systems via APIs presents security risks from untrusted access, requiring solutions like the Multi-Cloud Protocol, zero-trust models, and GraphQL to balance innovation with safeguards against data exposure and autonomous decision risks.

More MLOps.community episodes