
Hosted by VentureBeat
AI gets real here. On “Beyond the Pilot,” top business execs share what actually happens after the AI proof of concept - from infrastructure and org design to wins, failures, and ROI. Not theory, but deep dives into how they scaled AI that works.
32 episodes · publishes fortnightly · latest 2026-06-24 · ~32 min/episode
Rank
#61
Substance
85.0
/ 100
Breakdown
Scored 2026-07
Updated monthly
Across the index
#61 of 6182
Substance
Top 1%
outscores 99% of the index
Beyond The Pilot: Enterprise AI in Action ranks #61 on The B2B Podcast Index with a substance score of 85.0 out of 100, scored across 1 recent episode. It scores highest on guest caliber and specificity & evidence. Farhan is a senior practitioner actually running AI engineering at Shopify's scale, and his command of named internal systems (UDP, River, Tangle, Qwik, Sim Gym), specific toolchain decisions, and real infrastructure trade-offs confirms genuine operational depth rather than thought-leader abstraction.
Averaged across 1 recently scored episode, with cited evidence.
The episode delivers a genuine cluster of non-obvious operational insights - circuit breakers over token limits, River's mandatory-public-channel design creating unintended collaborative learning, Sim Gym for small merchants who lack AB testing traffic, and the Universal Distillation Platform pipeline - but dilutes them with extended agentic-commerce vision passages and standard platitudes about AI removing toil.
“we have circuit breakers in place which basically allow someone to get a message if something they're doing is long running and spending a lot of tokens”
“River only works in public, which means you can't go into a private channel and say, hey, river, like, help me build”
A few genuinely fresh frames - moving from 'AI reflexivity' to 'AI leverage,' the ideal AI Centaur being two humans plus an LLM rather than one, and the aspiration to let the distillation pipeline auto-select its own target model - but the episode also leans heavily on common Shopify infrastructure-first branding and the ubiquitous 'AI replaces tasks not jobs' line.
“we also moved in 2026 away from AI reflexivity to AI leverage”
“the ideal, like AI Centaur is not just human and LLM. It's like it could be two humans pairing, um, and the LLM helping you kind of remove the toil”
Farhan is a senior practitioner actually running AI engineering at Shopify's scale, and his command of named internal systems (UDP, River, Tangle, Qwik, Sim Gym), specific toolchain decisions, and real infrastructure trade-offs confirms genuine operational depth rather than thought-leader abstraction.
“we call it UDP Universal Distillation Platform. And what it allows us to do is you give it the teacher model, you give it data, you give it the evals and you give it the target model”
“we also have, uh, our own agentic platform like river, which again, switches between models”
Named internal platforms, specific model references (Qwen 3.5, Opus, GPT 5.2), a concrete 2x - 30x cost-reduction range, the 2021 GitHub Copilot deployment date, Toloka as a named data vendor, and the eval threshold illustration (70%) give this episode real evidential texture - though hard merchant-scale numbers and accuracy deltas are mostly absent.
“We, uh, deployed GitHub Copilot in 2021, which is a year before ChatGPT”
“we see savings, uh, in size from like 2x, which is huge, um, down sometimes down to like 30x”
The host asks a good set of follow-up probes - pushing on the frontier-vs-distilled production split, on whether traces could seed a proprietary coding model, and on who owns runaway token spend - but rarely challenges claims or creates productive friction, and several questions are open-ended scene-setters rather than sharp interrogations.
“What sort of split do you see at the moment between in production use of the Frontier LLMs versus distilled models?”
“Are you guys keeping the traces from your, you know, your developers with the aim of eventually being able to train your own coding model”
First period on the Index - history builds from here.
1 scored on substance · 32 tracked in total.
Add this badge to your site - it links back here and updates automatically as you rank.
<a href="https://index.fame.so/show/beyond-the-pilot-enterprise-ai-in-action" target="_blank" rel="noopener">
<img src="https://index.fame.so/badge/beyond-the-pilot-enterprise-ai-in-action/badge.svg" alt="Ranked #12 on The B2B Podcast Index" width="360" height="136" />
</a>Track Beyond The Pilot: Enterprise AI in Action's rank
Get an email whenever this show moves up or down the Index. Monthly at most, no spam.
Companies, products and tools that come up most across this show's episodes.
The themes that come up most across this show's episodes.
Podcasts that dig into the same topics.