Listen Score
LS 25
Global Rank
TOP 10%

ABOUT THIS PODCAST 🔗

Update frequency:
every 16 days
Average audio length:
47 minutes
Guest interviews
English
United States
52 episodes
since Feb. 7, 2024
episodic

LATEST EPISODE 🔗

We're hitting pause for the summer on new episodes of the Platform Engineering Podcast, but don’t worry, we’ve got some great episodes in the feed to keep you entertained in the meantime. If you’ve got suggestions for guests or topics, please reach out — I’m always eager to annoy people into joinin…

SEARCH PAST EPISODES

Search past episodes of Platform Engineering Podcast.

PREVIOUS EPISODES

Network calls fail in ways function calls never do - and once a monolith becomes microservices, reliability problems show up fast: retries amplify load, latency spikes cascade, and “what talks to what?” becomes hard to answer. William Morgan, co-creator of Linkerd and the person who coined “service…
When code gets cheaper to produce, feedback becomes the limiting factor - CI, reviews, and the handoffs between tools can quietly slow everything down. Rob Zuber breaks down what platform engineers are seeing as teams adopt AI-assisted development: more branch builds, new failure modes, and growing…
A lot of infrastructure and automation fails for ordinary reasons: rate limits, flaky networks, partial permissions, long-running jobs, and retries that vanish when the process restarts. Durable execution is a way to design systems that keep going anyway - without rebuilding a maze of queues, cron …
What happens when a non-deterministic AI system is asked to touch production telemetry or generate changes for an SRE pipeline? The cost of being “close enough” can be lost data, downtime, or a security incident. Cribl’s Nikhil Mungel joins Cory to break down what it takes to build AI that sysadmin…
When a flaky test can stall a merge queue, “just rerun CI” stops scaling fast. Cory talks with Trunk co-founder and CEO Eli Schleifer about the outer loop problems that show up as teams ship more code - especially with AI-assisted development increasing PR volume. They break down what a merge queue…
What happens when your “coworker” can generate code and changes faster than your team can review them, and production still has to stay up? William Collins breaks down what AI-Native Ops looks like when you take reliability seriously: where reasoning should stop, where deterministic automation shou…
Terraform drift, state wrangling, and a growing “tools for tools” stack are still daily work for many platform teams - despite a decade of DevOps talk and cloud maturity. Why does ops automation so often feel like it needs babysitting? Pavlo Baron breaks down where Infrastructure as Code tends to b…
Billions of requests a month on AWS Lambda can cost less than a single engineer’s laptop budget, but only if the architecture and developer workflow are designed for it. Justin Masse, Senior Platform DevOps Engineer at Extend, shares how Extend committed early to a serverless-first approach and bui…
What happens when nobody wrote the code running in your production environment? As AI-generated software becomes standard practice, platform engineers face a new challenge: operating systems without experts to consult. Nic Benders, Chief Technical Strategist at New Relic, has spent 15 years watchin…
Disclaimer: The podcast and artwork embedded on this page are from Cory O'Daniel, CEO of Massdriver, which is the property of its owner and not affiliated with or endorsed by Listen Notes, Inc.