ChatGPT and Beyond with Fexingo · 2026-07-13 · 9 min
Lucas and Luna explore the shift from training to inference as the economic bottleneck in AI. With NVIDIA up 3.4% in a week and AMD up 3.5%, they dig into why inference costs are now the crucial metric. Lucas explains how token prices are falling, but total inference demand is exploding - especially as agents and real-time applications scale. They discuss the 'inference gap' between what models can do and what deployment requires, and why this is reshaping chip design and cloud economics. A must-listen for anyone tracking where the real value in AI is moving. #AIInference #ComputeBottleneck #NVIDIA #AMD #TokenEconomics #AIAgents #LargeLanguageModels #GenerativeAI #Technology #TechStocks #AIInfrastructure #ChipDesign #CloudEconomics #DataCenters #FexingoBusiness #BusinessPodcast #ChatGPTandBeyond #ProductivityTools Keep every episode free: buymeacoffee.com/fexingo
Other episodes covering the same guests and topics, from across The B2B Podcast Index.