The B2B Podcast Index
Index
All categories
MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
MethodologySubmit
Best of:MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
An independent project byFame
SearchBest episodesGuestsInsightsMethodologySubmit a podcast
#427O11ycast77.0 / 100Get badge
← The Index
O11ycast artwork
Engineering & DevToolsNEW this period

O11ycast

Hosted by Heavybit

Exploring the observability side of software development.

90 episodes · publishes monthly · latest 2026-06-04 · ~38 min/episode

Rank

#427

Substance

77.0

/ 100

Breakdown

Scored 2026-07
Updated monthly

Engineering & DevTools rank

#49 of 289

Best B2B Engineering & DevTools Podcasts →

Across the index

#427 of 6183

Substance

Top 7%

outscores 93% of the index

Why it scores where it does

O11ycast ranks #427 on The B2B Podcast Index with a substance score of 77.0 out of 100, scored across 1 recent episode. It scores highest on guest caliber and insight density. Janaki is a hands-on engineer who built the production system under discussion, citing real internal workflows, specific error taxonomies, and named Amplitude products she shipped - solidly a practitioner guest, not a thought leader - though her seniority level is engineer rather than a senior decision-maker or founder, which limits the strategic depth on offer.

The five-dimension breakdown

Averaged across 1 recently scored episode, with cited evidence.

Insight Density

16.0 / 20

The episode contains a handful of genuinely useful ideas - compounding accuracy loss across agent steps, the primacy of 'the harness' over model upgrades, and the manual-tagging-to-auto-tagger pipeline - but these insights are spread thin across significant filler, repeated explanations, and a multi-minute detour into the guest's calendar art project that is irrelevant to any B2B operator.

“if each step is about 95% the way there, or 95% accurate, that leads to your response being Napkin math, about 60% accurate”

“a lot of the meat of how well our system does is related to, um, the harness around our agents rather than just upgrading the models”

Originality

14.0 / 20

The core framing - 'every failure becomes an eval' and treating manual tagging as seed data for automated evaluators - is practically grounded but already circulates widely in the AI engineering community; there are no genuinely contrarian or first-principles arguments, and the stone soup metaphor is borrowed from the guest's own pre-existing article rather than developed in the conversation.

“we parsed a lot of our internal notebooks and amplitude and generated nice tight evals from that that give us an input question”

“real investigations that took our own PM's hours to do then turned into evals that we were able to test our global agent system on”

Guest Caliber

17.0 / 20

Janaki is a hands-on engineer who built the production system under discussion, citing real internal workflows, specific error taxonomies, and named Amplitude products she shipped - solidly a practitioner guest, not a thought leader - though her seniority level is engineer rather than a senior decision-maker or founder, which limits the strategic depth on offer.

“I help build some of the systems that allow us to do analytics much faster”

“we just launched a product called Global Agent at Amplitude, which does the job of an analyst”

Specificity & Evidence

16.0 / 20

The episode delivers useful concrete detail - named products (Global Agent, Agent Analytics), a specific funnel-chart eval broken into discrete sub-checks (six steps, one-hour conversion window), and the 95%-per-step compounding math - but lacks hard business metrics such as user counts, error-rate improvements, or revenue impact, and the model naming is inconsistent ('Sana 4.5' vs Opus 4.6), undermining precision.

“Does the funnel chart now have the six steps that we're expecting for this particular storefront? Did it use the right conversion window of an hour?”

“if each step is about 95% the way there, or 95% accurate, that leads to your response being Napkin math, about 60% accurate”

Conversational Craft

14.0 / 20

Jessica asks several sharp, well-timed follow-ups ('How do you check your checker?', 'Is there a separate eval process pre-release?') and Ken usefully references the guest's written article, but neither host meaningfully pushes back on any claim, the art tangent consumes several minutes of irrelevant airtime, and the conversation never reaches productive disagreement or stress-tests the guest's framing.

“How do you check your checker?”

“Is there a separate eval process pre release for. We ask it the standard set of questions with this standard data set and see if it does a good job.”

Standout episodes

  • Ep. #91, Every Failure Becomes an Eval with Janaki Vivrekar

    2026-06-04

    77

Rank over time

First period on the Index - history builds from here.

Episodes

1 scored on substance · 60 tracked in total.

  • Ep. #91, Every Failure Becomes an Eval with Janaki Vivrekar

    2026-06-04 · 42 min

    77 / 100

Frequently asked

What is O11ycast's substance score?
O11ycast scores 77.0 out of 100 for substance and ranks #427 on The B2B Podcast Index. That puts it ahead of 93% of the B2B podcasts we rank and #49 of 289 in Engineering & DevTools. The score reflects insight density, originality, guest caliber, specificity and conversational craft across recent episodes - not downloads.
Is O11ycast worth listening to?
Yes - O11ycast outscores 93% of the B2B engineering & devtools podcasts and shows we rank on substance, so a engineering & devtools operator is likely to come away with something useful.
Who hosts O11ycast?
O11ycast is hosted by Heavybit.
How often does O11ycast publish?
O11ycast publishes monthly, has 90 episodes, released its most recent episode on 2026-06-04.
Which O11ycast episode should I start with?
Our highest-scoring recent episode is "Ep. #91, Every Failure Becomes an Eval with Janaki Vivrekar" (77/100) - a good place to start.

Show off your #49 rank in Engineering & DevTools

Add this badge to your site - it links back here and updates automatically as you rank.

Ranked #49 on The B2B Podcast Index
Embed code
<a href="https://index.fame.so/show/o11ycast" target="_blank" rel="noopener">
  <img src="https://index.fame.so/badge/o11ycast/badge.svg" alt="Ranked #49 on The B2B Podcast Index" width="360" height="136" />
</a>
Markdown & other formats →

Track O11ycast's rank

Get an email whenever this show moves up or down the Index. Monthly at most, no spam.

Listen / subscribe:WebsiteRSS

Frequently discusses

Companies, products and tools that come up most across this show's episodes.

AmplitudeGlobal AgentAgent AnalyticsHoneycombClaude Opus 4.6Claude Sonnet 4.5MCPSlackHeavybit

Guests who've appeared

Janaki Vivrekar

Topics this show covers

The themes that come up most across this show's episodes.

MCP (Model Context Protocol)Behavioral analyticsAmplitudeClaude Opus 4.6Global AgentClaude Sonos 4.5Agent Analyticsevaluation evalserror tagging taxonomyfunnel analysis

More Engineering & DevTools podcasts

See all →
  • Mik + One

    Dr. Mik Kersten: Author of Project to Product and Founder and CEO of Tasktop

    96.0
  • DevOps Daily with Fexingo

    Fexingo

    91.0
  • The Developer Tools Podcast with Fexingo

    Fexingo

    90.4
  • Rust in Production

    Matthias Endler

    90.0
  • mnemonic security podcast

    mnemonic

    88.8
  • Software Unscripted

    Richard Feldman

    88.0

Similar shows

Podcasts that dig into the same topics.

  • Masters of Automation

    Alp Uguray

    77.8
  • Embracing Marketing Mistakes

    Prohibition PR

    52.0
  • No Hacks

    Slobodan "Sani" Manić

    91.2
  • Platform Engineering Podcast

    Cory O'Daniel, CEO of Massdriver

    86.8
  • Cyber Sentries: AI Insight to Cloud Security

    TruStory FM

    86.4
  • AI Engineering Podcast

    Tobias Macey

    85.4