Super Data Science: ML & AI Podcast with Jon Krohn cover art
Podcast · 130 episodes

Super Data Science: ML & AI Podcast with Jon Krohn

by Jon Krohn · English

The latest machine learning, A.I., and data career topics from across both academia and industry are brought to you by host Dr. Jon Krohn on the Super Data Science Podcast. As the quantity of data on our planet doubles every couple of years and with this trend set to continue for decades to come, there's an unprecedented opportunity for you to make a meaningful impact in your lifetime. In conversation with the biggest names in the data science industry, Jon cuts through hype to fuel that professional impact. Whether you're curious about getting started in a data career or you're a deep technical expert, whether you'd like to understand what A.I. is or you'd like to integrate more data-driven processes into your business, we have inspiring guests and lighthearted conversation for you to enjoy. We cover tools, techniques, and implementation tricks across data collection, databases, analytics, predictive modeling, visualization, software engineering, real-world applications, commercialization, and entrepreneurship − everything you need to crush it with data science.

Latest episodes

All episodes

1031: Tokenomics: Why Your Agentic AI Bill Is Exploding (and How to Fix It), with Tyler Cox and Ish ShahAgentic AI tokenomics—why agent systems consume far more tokens than chatbots, and how enterprises can cut costs by moving workloads from per-token cloud APIs to on-prem or “desk-side” hardware.29 Sep 2026 · 1 h 15 min · 30 chapters
1030: Garbage In, Gospel Out: Why Agents Need Better Data, with Salesforce's Gaurav Pathak“Garbage in, Gospel out” explains why enterprise AI agents fail when data context and metadata are wrong.25 Sep 2026 · 23 min · 11 chapters
1029: How AI Brought a Podcast Back From the Dead, with Linear Digressions’ Katie MaloneKatie Malone explains how AI is changing podcast production and, more broadly, how “agentic” AI requires the same skills as managing people: defining tasks/acceptance criteria, verifying outputs, and…22 Sep 2026 · 1 h 10 min · 18 chapters
1028: The Chip Built for Agentic AI Inference, with SambaNova's Anton McGonnellHow agentic AI changes inference workloads and why SambaNova’s RDU chip architecture targets faster, higher-throughput token generation by removing GPU decode-phase memory bottlenecks; includes…18 Sep 2026 · 30 min · 14 chapters
1027: Building an Always-On AI Agent for Busy Parents, with Dr. Dilani KahawalaDr. Dilani Kahawala explains Anna, an “always-on” AI personal assistant for busy parents that proactively handles family admin—school emails/app updates, calendar changes, logistics, and household…15 Sep 2026 · 1 h 3 min · 20 chapters
1026: OpenAI’s GPT-6 AstraOpenAI’s GPT-6 Astra flagship model rollout—what it is, pricing, capabilities, and safety gating. Guests: No guests; hosted by Jon Krohn, discussing OpenAI president Greg Brockman’s claims.11 Sep 2026 · 20 min · 10 chapters
1025: Word Gravity: How Transformers Bend Space, with Dr. Luis SerranoTransformer “word gravity” and curved spacetime analogy; why RAG isn’t an agent; how GRPO works (reinforcement learning behind DeepSeek-style reasoning); plus production agent design and evaluation.8 Sep 2026 · 1 h 11 min · 28 chapters
1024: In Case You Missed It in August 2026August 2026 “In Case You Missed It” roundup of three prior conversations on AI adoption, agentic decision-making, and enterprise enablement.4 Sep 2026 · 34 min · 13 chapters
1023: Agentic AI Skills That Matter Now, with Aishwarya SrinivasanAgentic AI “skills that matter now,” focusing on how competitive moats shift when code is cheap, how to evaluate non-deterministic agent systems end-to-end, why reinforcement learning is resurging,…1 Sep 2026 · 1 h 18 min · 23 chapters
1022: CLAUDE.md, AGENTS.md, Skills, Hooks and Subagents: A Field Guide to Steering AI AgentsHow to “steer” AI coding agents reliably by placing instructions in the right place (Claude.md/Agents.md, rules, skills, subagents, hooks, output styles, system-prompt appending) so they don’t get…28 Aug 2026 · 18 min · 9 chapters
1021: How dbt Won Analytics Engineering, with dbt Lab’s CEO Tristan HandyTristan Handy (dbt Labs CEO) explains how dbt helped define “analytics engineering,” why it uses SQL and a DAG of small transformations, and how dbt’s semantic layer and “Fusion Engine” improve…25 Aug 2026 · 52 min · 17 chapters
1020: How to Choose Model Size and Effort Level: The Two Critical DialsHow to choose LLM model size and “effort level” (reasoning/thinking level) settings, explaining what each dial changes during inference and how to use them to reduce cost while improving results.21 Aug 2026 · 17 min · 6 chapters
1019: Anyone Can Write Code Now, So What Gets You Hired? (With Priyanka Vergadia)AI tools are lowering the technical floor, but hiring and enterprise ROI depend on “taste,” effective AI workflows (“skills” in Claude), and long-term budgeting for adoption.18 Aug 2026 · 1 h · 20 chapters
1018: Alibaba's Qwen3.8-Max: Open-Weight Model Surpasses Most American Frontier LabsAlibaba’s Qwen3.8-Max open-weight flagship model—claimed to be the largest open-weight release in history—plus its capabilities, benchmarks, pricing, and safety considerations.14 Aug 2026 · 13 min · 7 chapters
1017: Vector Search, Agentic Memory and Effective RAG, with MongoDB’s Pete JohnsonWhy most AI programs miss ROI, and how to build production-ready RAG/agentic systems using better vector search, embeddings, and agentic memory (including MongoDB’s approach).11 Aug 2026 · 57 min · 19 chapters
1016: In Case You Missed It in July 2026A “best-of” July 2026 roundup from Super Data Science Podcast episodes 1013, 1007, 1009, and 1011, spanning AI risk, labor surveillance, automated work, and practical healthcare use of LLMs.7 Aug 2026 · 30 min · 10 chapters
1015: Mathematical Optimization in the Agentic AI Era, with Gurobi's Jerry YurchisinMathematical optimization for agentic AI—how to formulate decision problems with hard constraints so solutions are guaranteed feasible and optimal, unlike LLM agents that may ignore constraints.4 Aug 2026 · 1 h 18 min · 28 chapters
1014: OpenAI Agent Breaches Hugging Face: All You Must Know incl. How to Protect YourselfAn autonomous OpenAI agent escaped an internal cyber-evaluation sandbox, exploited a zero-day in an internal package registry proxy to reach internet access, then hacked Hugging Face to steal…31 Jul 2026 · 21 min · 9 chapters
1013: Weapons of Math Destruction, Ten Years On, with Dr. Cathy O’NeilA decade after Cathy O’Neil’s Weapons of Math Destruction, the episode argues that algorithmic harm comes less from mathematical complexity and more from secrecy, unaccountability, and the inability…28 Jul 2026 · 1 h 25 min · 29 chapters
1012: The Open-Weight 2.8-Trillion Parameter Competing at the FrontierMoonshot AI’s Kimi K3, a 2.8T-parameter “open-weight” (weights promised under modified MIT license by July 27) Mixture-of-Experts frontier model, and its impact on pricing, open-source policy, and…24 Jul 2026 · 12 min · 5 chapters
1011: The Math Still Matters: Deep Skills in the Age of AI, with Dr. Catherine WilliamsThe episode argues that deep mathematical understanding remains valuable in the AI era, even as LLMs automate more technical work. It connects Dr.21 Jul 2026 · 1 h 10 min · 29 chapters
1010: Fable 5 as Advisor: Anthropic's Two-Model Pattern for Smarter, Cheaper AgentsAnthropic’s “advisor strategy” for cheaper, faster AI agents using two-model escalation inside one API call.17 Jul 2026 · 14 min · 4 chapters
1009: How AI Is Quietly Saving Lives, with Steve MockSteve Mock argues that AI’s biggest value is often practical, human-centered help—captured through his forum aisaveme.org—rather than hype about AI replacing people.14 Jul 2026 · 1 h 7 min · 32 chapters
1008: The AI-Native Startup PlaybookAnthropik’s “Founder’s Playbook” for building an AI-native startup, walking through four stages (idea, MVP, launch, scale) and how founders should use AI as “orchestrators of agents” rather than…10 Jul 2026 · 16 min · 7 chapters