Episódios

  • 820: OpenAI's o1 "Strawberry" Models
    Sep 20 2024
    Jon Krohn takes OpenAI’s new models (o1-preview and o1-mini) for a spin in this Five-Minute Friday, learning their key strengths and limitations, and how the o1 series may represent yet another landmark for generative AI. Additional materials: www.superdatascience.com/820  Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.
    Exibir mais Exibir menos
    27 minutos
  • 819: PyTorch: From Zero to Hero, with Luka Anicin
    Sep 17 2024
    SuperDataScience veteran and Udemy teacher Luka Anicin is on the podcast to talk about his brand-new course, “PyTorch: From Zero to Hero”, available exclusively on superdatascience.com. Host Jon Krohn asks Luka why he feels that every data scientist should consider PyTorch as their default Python library, and why “keeping it simple” can secure the success of a machine learning project. This episode is brought to you by AWS Inferentia and AWS Trainium, and by Gurobi, the Decision Intelligence Leader. Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: • About the PyTorch library [03:29] • Why PyTorch became so popular [25:24] • How to increase accuracy and efficiency in PyTorch [31:49] • How to utilize transfer learning [35:44] • Why real-world projects are essential to data scientists [41:10] • About Datablooz [46:49] Additional materials: www.superdatascience.com/819
    Exibir mais Exibir menos
    1 hora e 6 minutos
  • 818: In Case You Missed It in August 2024
    Sep 13 2024
    Experts from AI and data science discuss the impact and benefits of decentralization, the importance of structuring AI systems in business, and why knowing the basics will always matter for data engineers. Listen to Shingai Manjengwa (episode 809), Daniel Hulme (episode 807), Jerry Yurchisin (episode 813) and Nick Elprin (episode 811) explore a future world of work that rewards continuing learners, sets tasks for the people best suited to complete them rather than those whose job titles reflect the spec, and applies a fleet of ‘AI agents’ to solve complex business tasks. Additional materials: www.superdatascience.com/818  Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.
    Exibir mais Exibir menos
    30 minutos
  • 817: The Positron IDE, Tidy NLP and MLOps with Dr. Julia Silge
    Sep 10 2024
    Dr. Julia Silge, Engineering Manager at Posit, introduces the brand-new Positron IDE, perfect for exploratory data analysis and visualization. She also lays out her top picks for LLMs that boost coding efficiency and discusses when traditional NLP methods might be the smarter choice over LLMs. Plus, Julia highlights some must-know open-source libraries that make managing MLOps easier than ever. Tune in for insights that every data scientist, ML engineer, and developer will find useful. This episode is brought to you by Gurobi, the Decision Intelligence Leader, and by ODSC, the Open Data Science Conference. Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: • Overview of Posit and Positron IDE [05:20] • How the needs of a data scientist differ from those of a software developer [10:54] • How to contribute to the open-source Positron [19:50] • MLOps and Vetiver: Tools for deploying and maintaining ML models [37:01] • Natural Language Processing (NLP) and the Tidyverse approach [50:34] • The role of AI and LLMs in data science education [1:24:18] Additional materials: www.superdatascience.com/817
    Exibir mais Exibir menos
    1 hora e 36 minutos
  • 816: Explaining AGI to a 94-Year-Old
    Sep 6 2024
    Jon Krohn takes on a listener's challenge to explain his work in data science to his 94-year-old grandmother, Annie. This heartwarming conversation covers what data is, the role of a data scientist, and breaks down artificial intelligence (AI) and artificial general intelligence (AGI) in simple terms. The episode provides a fresh take on how to communicate complex topics to a lay audience, offering both clarity and insight. Additional materials: www.superdatascience.com/816  Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.
    Exibir mais Exibir menos
    20 minutos
  • 815: Polars: Faster DataFrame Ops, with Marco Gorelli
    Sep 3 2024
    Polars, Python, Narwhals, Rust, and Pandas: Marco Gorelli talks to Jon Krohn about the many ways to use the newest data libraries available, the joys of open-source development, and the best method to win prizes in forecasting competitions. This episode is brought to you by AWS Inferentia and AWS Trainium, by Babbel, the science-backed language-learning platform, and by Gurobi, the Decision Intelligence Leader. Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: • When to use Polars vs Pandas [08:26] • How Polars optimizes string operations and data processing [20:08] • Where Narwhals outstrips Polars and Pandas [48:37] • The benefits of using Altair [55:21] • Addressing the lack of women in data science [1:09:58] • How to win a forecasting competition [1:16:58] Additional materials: www.superdatascience.com/815
    Exibir mais Exibir menos
    1 hora e 27 minutos
  • 814: Summer Reflections
    Aug 30 2024
    As summer winds down, this episode shifts focus from the usual tech discussions to something more personal: reflecting on the importance of balancing work with life’s simple pleasures. While the world of data science and AI continues to evolve rapidly, it's essential to remember that true success isn't just about professional milestones. It’s also about cherishing the moments that make life meaningful. Tune in for a brief but impactful reflection on how to redefine success to include not just achievements, but also the everyday joys that often go unnoticed. Additional materials: www.superdatascience.com/814  Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.
    Exibir mais Exibir menos
    4 minutos
  • 813: Solving Business Problems Optimally with Data, with Jerry Yurchisin
    Aug 27 2024
    Jerry Yurchisin from Gurobi joins Jon Krohn to break down mathematical optimization, showing why it often outshines machine learning for real-world challenges. Find out how innovations like NVIDIA’s latest CPUs are speeding up solutions to problems like the Traveling Salesman in seconds. Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information. In this episode you will learn: • The Burrito Optimization Game and mathematical optimization use cases [03:36] • Key differences between machine learning and mathematical optimization [05:45] • How mathematical optimization is ideal for real-world constraints [13:50] • Gurobi’s APIs and the ease of integrating them [21:33] • How LLMs like GPT-4 can help with optimization problems [39:39] • Why integer variables are so complex to model [01:02:37] • NP-hard problems [01:11:01] • The history of optimization and its early applications [01:26:23] Additional materials: www.superdatascience.com/813
    Exibir mais Exibir menos
    1 hora e 44 minutos