
Hosted by Cory O'Daniel, CEO of Massdriver
The Platform Engineering Podcast is a show about the real work of building and running internal platforms - hosted by Cory O’Daniel, longtime infrastructure and software engineer, and CEO/cofounder of Massdriver.
51 episodes · publishes fortnightly · latest 2026-06-24 · ~48 min/episode
Rank
#36
Substance
86.8
/ 100
Breakdown
Scored 2026-07
Updated monthly
Across the index
#36 of 6182
Substance
Top 1%
outscores 99% of the index
Platform Engineering Podcast ranks #36 on The B2B Podcast Index with a substance score of 86.8 out of 100, scored across 5 recent episodes. It scores highest on guest caliber and insight density. Cornelia Davis is an exceptionally credible guest: 30+ years in distributed systems, 7 years as VP Technology at Pivotal (architect of Cloud Foundry), author of 'Cloud Native Patterns' (10,000+ copies sold), now Principal Technologist at Temporal with deep domain authority. She demonstrates genuine practitioner experience with real failure scenarios, design tradeoffs, and customer deployments. She speaks with earned skepticism (e.g., on LLM nondeterminism, the 60-80% failure-handling tax) and avoids vapid pontification. This is a rare instance of a guest who has actually built and shipped at scale across multiple eras of infrastructure.
Averaged across 5 recently scored episodes, with cited evidence.
The episode densely packs substantive distributed systems concepts applied to infrastructure orchestration. Cornelia Davis articulates core durable execution mechanics (durable retries, state management, task queues, entity workflows/digital twins) with concrete platform engineering examples (Terraform composability, quota failures, orphaned infrastructure). However, some sections devolve into promotional asides and repetition (explaining durable execution multiple times, brief AI tangent), and certain explanations could be more granular. The quota/IAM error scenario discussion is rich but somewhat incomplete on implementation details.
“sixty to eighty percent of the code that is built for applications and distributed systems is failure handling code”
“durable execution is a programming model that allows you to write your code as if those failures didn't exist”
Davis articulates a genuinely fresh angle: reframing platform engineering as an unsolved distributed systems problem and applying durable execution patterns from financial/backend systems to infrastructure orchestration. The digital twin framing for long-lived infrastructure state is conceptually novel. However, the core durable execution model itself (event sourcing, task queues, retry policies) is not new - it's acknowledged as financial-systems heritage. The episode lacks contrarian or first-principles challenges; it mostly advocates for adopting proven patterns rather than questioning assumptions. The AI integration section is shallow and somewhat opportunistic.
“what we do in the platform space is a distributed systems problem...we still haven't brought [those patterns] over into the platform space yet”
“I actually prefer a different term...the notion of a digital twin...you always have a logical and digital analog to this very real thing”
Cornelia Davis is an exceptionally credible guest: 30+ years in distributed systems, 7 years as VP Technology at Pivotal (architect of Cloud Foundry), author of 'Cloud Native Patterns' (10,000+ copies sold), now Principal Technologist at Temporal with deep domain authority. She demonstrates genuine practitioner experience with real failure scenarios, design tradeoffs, and customer deployments. She speaks with earned skepticism (e.g., on LLM nondeterminism, the 60-80% failure-handling tax) and avoids vapid pontification. This is a rare instance of a guest who has actually built and shipped at scale across multiple eras of infrastructure.
“I spent seven years at Pivotal as the VP of Technology, where she helped shape Cloud Foundry”
“she's also the author of 'Cloud Native Patterns' and she spent more than three decades helping developers build resilient distributed systems”
The episode contains concrete examples (Terraform resource decomposition, credential vs. quota vs. IAM permission failures, Redis cluster boot delays, six nines SaaS availability) and specific tools (Temporal SDK decorators, task queues, event sourcing). However, specificity is inconsistent: the Terraform/IAM error-handling scenario is well-articulated but lacks actual code; the AI migration use case (3 weeks vs. 6 months) cites a named customer conference talk but provides no quantitative breakdown; Redis boot-time claim ("an hour and a half") is anecdotal; LLM token-burn economics mentioned but not quantified. Digital twin pattern is explained conceptually but no concrete deployment metrics or observability output examples given.
“Redis...it takes an hour and a half to boot up the Redis cluster”
“They had scheduled six months for the migration from some of their legacy applications...They did it in three weeks”
Cory O'Daniel asks sharp, scenario-grounded questions that expose nuance (quota vs. IAM vs. credential failures; pause/resume workflows; dead-letter queue handling; AI nondeterminism). He follows up meaningfully and pushes back (e.g., "How do you decorate for that scenario?" when error modes conflict). Davis largely engages substantively. However, the host occasionally allows lengthy monologues without tighter follow-up, misses opportunities to challenge (e.g., doesn't press on whether Temporal's zero-resource wait state claim is universal across platforms, or probe the AI migration case deeper). The conversation also includes two promotional segments (Massdriver ad, Temporal SaaS pitch) that interrupt flow. Overall solid but not as probing as it could be.
“So in scenarios like that where it's like...the quota one, it's not retryable but it also is? You know what I mean.”
“Are those Temporal skills also open source and can be used with the open source Temporal?”
First period on the Index - history builds from here.
10 scored on substance · 51 tracked in total.
What Do Service Meshes Actually Solve? (William Morgan, Buoyant/Linkerd)
2026-06-24 · 56 min
Continuous Integration at Agentic Velocity with CircleCI’s Rob Zuber
2026-06-10 · 50 min
Durable Execution for Real‑World Failures with Temporal’s Cornelia Davis
2026-05-27 · 46 min
You Need AI Sysadmins Can Trust, With Cribl's Nikhil Mungel
2026-05-13 · 55 min
Green CI and Merge Queue Mastery with Trunk’s Eli Schleifer
2026-04-15 · 50 min
AI-Native Ops: Making AI Safe for Production with William Collins
2026-04-01 · 1h 3m
Infrastructure as Code's Hidden Problem with Pavlo Baron
2026-03-18 · 58 min
Why Extend Went All-In on Serverless Platform Engineering
2026-03-04 · 1h 2m
Observability in the AI Era with New Relic's Nic Benders
2026-02-18 · 51 min
Simplicity at Scale: Cleaning House for Platform Teams with Brian Childress
2025-12-17 · 41 min
Add this badge to your site - it links back here and updates automatically as you rank.
<a href="https://index.fame.so/show/platform-engineering-podcast" target="_blank" rel="noopener">
<img src="https://index.fame.so/badge/platform-engineering-podcast/badge.svg" alt="Ranked #8 on The B2B Podcast Index" width="360" height="136" />
</a>Track Platform Engineering Podcast's rank
Get an email whenever this show moves up or down the Index. Monthly at most, no spam.
The themes that come up most across this show's episodes.
Podcasts that dig into the same topics.