The B2B Podcast Index
Index
All categories
MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
MethodologySubmit
Best of:MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
An independent project byFame
SearchBest episodesGuestsInsightsMethodologySubmit a podcast
Index/Engineering & DevTools/TestGuild News Show
TestGuild News Show artwork

The AI Testing Trust Crisis: Verification Costs, Gamed Benchmarks, and What Comes Next TGNS186

TestGuild News Show · 2026-06-01 · 10 min

0:00--:--

Episode notes

Have you seen the new testing tool that claims to give you fully working end-to-end tests in five minutes with zero setup? What are some of the ways AI agents are quietly gaming their own benchmarks, and what does that mean for how you evaluate them? How do you keep test-driven development alive when AI is the one writing the code? Find out in this episode of the TestGuild News Show for the week of June 1st. So, grab your favorite cup of coffee or tea, and let's do this. Time Item URL 0:00 Intro 0:24 Testifly 1:13 AI False Confident principle 2:46 Webinar of the Week 3:38 AI Agent Cheating 4:44 TDD for AI 6:10 Webwright 7:29 AI Quality Manifesto 8:45 Claude Workflows

More from TestGuild News Show

All episodes →
  • New Playwright, LinkedIn's AI Tester, AI for Selenium and More TGNS18946 / 100
  • Is AI Coming for Testers, or Are You About to Win Big? TGNS188
  • From Vibe Slop to AgentOps, Postman AI, Bug Report to Release Sign Off and More! TGNS187
  • MCP Servers, Microcks, and the New AI Testing Stack TGNS185
  • Testing in the Age of AI: What's Working, What's Not, and What's Next TGNS184
Explore the best B2B Engineering & DevTools podcasts →
All TestGuild News Show episodes →