The B2B Podcast Index
Index
All categories
MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
MethodologySubmit
Best of:MarketingSalesSaaSFinanceHROpsLeadershipCustomer SuccessAI & DataProductStartups & FoundersRevOpsEngineering & DevTools
An independent project byFame
SearchBest episodesGuestsInsightsMethodologySubmit a podcast
Index/AI & Data/Data Brew by Databricks
Data Brew by Databricks artwork

Benchmarking Domain Intelligence | Data Brew | Episode 45

Data Brew by Databricks · 2025-04-24 · 32 min

0:00--:--

Episode notes

In this episode, Pallavi Koppol, Research Scientist at Databricks, explores the importance of domain-specific intelligence in large language models (LLMs). She discusses how enterprises need models tailored to their unique jargon, data, and tasks rather than relying solely on general benchmarks. Highlights include: - Why benchmarking LLMs for domain-specific tasks is critical for enterprise AI. - An introduction to the Databricks Intelligence Benchmarking Suite (DIBS). - Evaluating models on real-world applications like RAG, text-to-JSON, and function calling. - The evolving landscape of open-source vs. closed-source LLMs. - How industry and academia can collaborate to improve AI benchmarking.

More from Data Brew by Databricks

All episodes →
  • Reinforcement Fine-Tuning and the Future of Specialized AI Models76 / 100
  • SWE-bench & SWE-agent | Data Brew | Episode 44
  • Enterprise AI: Research to Product | Data Brew | Episode 43
  • Multimodal AI | Data Brew | Episode 42
  • Age of Agents | Data Brew | Episode 41
Explore the best B2B AI & Data podcasts →
All Data Brew by Databricks episodes →