The B2B Podcast Index
Index
All categories
MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
MethodologySubmit
Best of:MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
An independent project byFame
SearchBest episodesGuestsInsightsMethodologySubmit a podcast
Index/Engineering & DevTools/cloud2030
cloud2030 artwork

Back After a Break

cloud2030 · 2026-05-29 · 29 min

0:00--:--

Episode notes

In this episode, we discuss the rising cost of using AI and how usage-based pricing, model changes, and capacity limits are affecting daily work as AI moves from experimentation into operational use. We also talk about multi-model workflows, hybrid infrastructure, and examples of using hosted models alongside open models locally for tasks such as writing and named entity resolution. We get into the need for enterprises to run their own AI infrastructure, including questions around GPU pooling, routing, reservation, data sovereignty, and service levels. Transcript:

More from cloud2030

All episodes →
  • AI UX Building60 / 100
  • Vibe Coding for Ops Project Restart [TechOps]
  • Kubernetes as Common Platform
  • Kubecon SC25 Debrief
  • AWS Outage
Explore the best B2B Engineering & DevTools podcasts →
All cloud2030 episodes →