Super Data Science: ML & AI Podcast with Jon Krohn

Jon Krohn

The latest machine learning, A.I., and data career topics from across both academia and industry are brought to you by host Dr. Jon Krohn on the Super Data Science Podcast. As the quantity of data on our planet doubles every couple of years and with this trend set to continue for decades to come, there's an unprecedented opportunity for you to make a meaningful impact in your lifetime. In conversation with the biggest names in the data science industry, Jon cuts through hype to fuel that professional impact. Whether you're curious about getting started in a data career or you're a deep technical expert, whether you'd like to understand what A.I. is or you'd like to integrate more data-driven processes into your business, we have inspiring guests and lighthearted conversation for you to enjoy. We cover tools, techniques, and implementation tricks across data collection, databases, analytics, predictive modeling, visualization, software engineering, real-world applications, commercialization, and entrepreneurship − everything you need to crush it with data science.

  1. 14h ago

    1021: How dbt Won Analytics Engineering, with dbt Lab’s CEO Tristan Handy

    In Episode #1021, Tristan Handy (Founder and CEO of dbt Labs) joins Jon Krohn to explain how a study of about a hundred companies in 2016 became analytics engineering, and then became a tool that over a hundred thousand data teams rely on. Tristan coined the term, chose SQL when Spark was the fashionable answer, and spent a decade turning down acquisition offers because none of them were good for the people using dbt. He is now merging dbt Labs with Fivetran and taking on the presidency of the combined company, the first deal he says cleared that bar. In this episode, Tristan walks through what dbt does to your raw data, argues that the semantic layer matters more once analytics agents are asking the questions, explains the type safety behind the Fusion engine, and details how a 12-kilobyte skill file collapses a million-dollar migration into six weeks. Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://www.superdatascience.com/1021⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: (00:07:22) Why Tristan chose SQL over Spark, and what progressive complexity means (00:10:44) How a dbt project turns raw data into modeled tables (00:17:50) Why a decade of acquisition offers kept failing his one test (00:40:47) How 12-kilobyte skill files cut year-long migrations to six weeks

    1021: How dbt Won Analytics Engineering, with dbt Lab’s CEO Tristan Handy
  2. Aug 18

    1019: Anyone Can Write Code Now, So What Gets You Hired? (With Priyanka Vergadia)

    In Episode #1019, Priyanka Vergadia (founder of The Cloud Girl, former Senior Director of AI Transformation at Microsoft and Head of North America Developer Relations at Google) joins Jon Krohn to explain why almost every company has bought AI tools and almost none of them are seeing a return. Her fix is a budget split that will make any CFO wince: seven dollars on training employees for every dollar spent on the tools themselves. Having spent a decade turning dense cloud and AI concepts into sketches that a quarter-million developers actually remember, and having carried GitHub Copilot into Fortune 100 boardrooms, she has watched the gap between tool purchase and real production use up close. In this episode, Priyanka defines the elusive quality she calls taste, walks through how she structures Claude skills so her output stops being slop, unpacks her 10-20-70 framework, and shares breaking news about what she is building next. Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://www.superdatascience.com/1019⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: (00:10:39) What “taste” actually means and why Priyanka now interviews for it (00:31:39) How to build a Claude skill by breaking a task into explicit sub-tasks (00:36:11) The 10-20-70 framework for AI budgets (00:47:52) The weekend exercise for finding what makes you different

    1019: Anyone Can Write Code Now, So What Gets You Hired? (With Priyanka Vergadia)
  3. Aug 11

    1017: Vector Search, Agentic Memory and Effective RAG, with MongoDB’s Pete Johnson

    In Episode #1017, Pete Johnson (Field CTO of AI at MongoDB) joins Jon Krohn to explain why four out of five organizations have AI steering committees and success metrics, yet only one in five sees a return on the investment. Having made nineteen stops across six countries this year advising more than a hundred companies on their AI strategies, Pete has an unusually wide view of what is actually working in production. In this episode, he traces the history of SQL and denormalization, unpacks why the embedding model is the most underrated choice in a RAG pipeline, explains Matryoshka embeddings and lays out what better agentic memory looks like. Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://www.superdatascience.com/1017⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: (00:06:34) Why the AI ROI gap happens and what to do differently (00:18:21) Jevons paradox, bank tellers and toll booth workers (00:24:05) From Codd’s 1970 paper to denormalization (00:32:31) Why the embedding model is not a commodity (00:40:20) What better agentic memory looks like

    1017: Vector Search, Agentic Memory and Effective RAG, with MongoDB’s Pete Johnson
  4. Aug 4

    1015: Mathematical Optimization in the Agentic AI Era, with Gurobi's Jerry Yurchisin

    In Episode #1015, Jerry Yurchisin (manager of decision intelligence strategy at Gurobi Optimization) joins Jon Krohn to explain the AI technology that makes breaking a constraint mathematically impossible. Large language models will confidently claim they've optimized your business while ignoring the one constraint that could cost millions, whereas optimization treats constraints as hard guarantees. Jerry lays out the division of labor he sees for the agentic era: agents help you frame the problem, write the formulation and generate the code, then hand off to a solver like Gurobi, soon callable via MCP servers. In this episode, Jerry breaks down the three building blocks of any optimization model, traces the leap in non-linear solving, explains how to pitch optimization to your CFO and to the planners whose jobs it touches, and shares case studies spanning energy grids, retirement planning and USA Cycling's Paris 2024 gold. Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://www.superdatascience.com/1015⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: (02:42) The three building blocks of an optimization model (21:43) Where optimization fits in the agentic AI era (29:58) Inside the Gurobi Intelligence Hub (39:40) Energy, retirement planning and a cycling gold medal (50:58) How to sell optimization inside your organization

    1015: Mathematical Optimization in the Agentic AI Era, with Gurobi's Jerry Yurchisin
4.6
out of 5
305 Ratings

About

The latest machine learning, A.I., and data career topics from across both academia and industry are brought to you by host Dr. Jon Krohn on the Super Data Science Podcast. As the quantity of data on our planet doubles every couple of years and with this trend set to continue for decades to come, there's an unprecedented opportunity for you to make a meaningful impact in your lifetime. In conversation with the biggest names in the data science industry, Jon cuts through hype to fuel that professional impact. Whether you're curious about getting started in a data career or you're a deep technical expert, whether you'd like to understand what A.I. is or you'd like to integrate more data-driven processes into your business, we have inspiring guests and lighthearted conversation for you to enjoy. We cover tools, techniques, and implementation tricks across data collection, databases, analytics, predictive modeling, visualization, software engineering, real-world applications, commercialization, and entrepreneurship − everything you need to crush it with data science.

You Might Also Like