Skip to content
Artwork for It's About Data

It's About Data

Matthew Housley and Tony Baer

We discuss data, artificial intelligence, and whatever else we feel like talking about. Cohosted by Tony Baer and Matt Housley. Also available on YouTube:
https://www.youtube.com/channel/UC6XFbzM-Vn1ztmEQFDVzYCQ

Play
  • 20 episodes
  • Avg 20 min
  • English
  • S1 · E93
    Yesterday · 20 min

    The State of Enterprise AI (Rob Strechay—VentureBeat)

    Use the code MATT20 for a 20% discount on a one year subscription to The O'Reilly Learning Platform. https://learning.oreilly.com/signup/ Your subscriptions support this podcast. Rob Strechay (Lead Analyst, VentureBeat) joins us to discuss the state of enterprise AI. Is the term context graph useful? What role will open weight models play in the enterprise? How are enterprises managing AI costs?

  • S1 · E92
    Tuesday · 20 min

    The Year of The Graph

    Are we finally going to experience the year of the graph database? According to Arthur Bigeard, (founder gdotv.com) graph databases are having a moment driven by gen AI and graph on relational tools. We also talk about the evolution of data development, graph visualization, ontologies, and socialization of graph technology. Fundamentals of Data Engineering is available through the O'Reilly Learning Platform along with over 60K other titles. Sign up for a free trial here. Get a 20% discount on a one year O'Reilly Learning Platform subscription using the code MATT20. https://learning.oreilly.com/signup/ Your subscriptions support this podcast.#graphdatabase #genai

  • S1 · E91
    Monday · 27 min

    Nvidia/Hugging Face

    We're joined by Miriah Peterson and Matt Sharp to discuss Nvidia's acquisition of Hugging Face. Coverage from The Information Domesticating AI, Matt and Miriah's podcast with Chris Brousseau. Hugging Face: huggingface.co Matt Sharp's book LLMs in Production is available through the O'Reilly Learning Platform along with over 60K other titles.Sign up for a free trial here. Get a 20% discount on a one year O'Reilly Learning Platform subscription using the code MATT20. https://learning.oreilly.com/signup/ Your subscriptions support this podcast.

  • S1 · E90
    August 26 · 24 min

    AWS Acquires DuckLabs

    This is a breaking news episode responding to the announcement that Amazon has signed an agreement to acquire DuckLabs, the developer of the DuckDB analytics database. The announcement from DuckLabs: https://ducklabs.com/ducklabs-to-join-aws Joe Reis and Matt Housley interviewing Jordan Tigani about the founding of MotherDuck: https://www.youtube.com/watch?v=E2fi-Y6RiTw Our interview with Nikita Shamgunov on Neon and Lakebase. (Ep 73). https://youtu.be/zpqM-YWrXB4?si=Ik-0zjrjT1_mU-Rz. MotherDuck acquires Tower Computing: https://motherduck.com/blog/motherduck-acquires-tower/ #duckdb #dataanalytics

  • S1 · E88
    August 11 · 9 min

    Business Intelligence with LLMs (guest Ryan Dolley)

    Ryan Dolley joins the show to discuss his work on an upcoming book. This episode was recorded at DataTune 2026 (https://datatuneconf.com/).#businessintelligence #largelanguagemodels

  • S1 · E87
    July 28 · 19 min

    The Rapidly Evolving Job Market (with Stephen Messer)

    Stephen Messer (cofounder of Collective[i]) joins us to discuss an AI-driven talent shortage, the future of education, non-generative AI and the business prospects of foundation model developers.This episode is a continuation of our conversation in episode 86.https://spotifycreators-web.app.link/e/vI0LwVLgY4b#artificialintelligence

  • S1 · E84
    July 8 · 16 min

    Managing the Meaning of Data (Dipti Borkar—Microsoft)

    This is a continuation of our conversation with Dipti Borkar from episode 81. We discuss Microsoft Fabric IQ as a tool for managing meaning in data, the problem of data interop, and the rapidly evolving role of data engineers.Episode 81:https://spotifycreators-web.app.link/e/K7ACFfYNq4b

  • S1 · E83
    July 6 · 20 min

    Harnessing Unstructured Data Part II (with Kevin Petrie and Merv Adrian)

    This is a continuation of our discussion in episode 82. Kevin explains that AI agents are now a "fox in charge of the hen house," but this actually portends great things in the longstanding quest for data quality.Harnessing Unstructured Data for AI Innovation (BARC Study):https://barc.com/research/harnessing-unstructured-data-for-ai-innovation/Episode 82, where we introduce the BARC study:https://open.spotify.com/episode/5TexEQAYm2S5cOWUvXGgY1?si=88d763ba0ae045a9​

  • S1 · E82
    July 2 · 18 min

    Harnessing Unstructured Data for AI Innovation (with Kevin Petrie and Merv Adrian)

    Kevin Petrie and Merv Adrian (BARC) join us to discuss a recent study:"Harnessing Unstructured Data for AI Innovation: Problems, Practices, and Principles for Success"This episode is part of a two part series. In the first part, we discuss the findings of the study.The study is available here:https://barc.com/research/harnessing-unstructured-data-for-ai-innovation/

  • S1 · E80
    June 27 · 19 min

    Moving Up the Data Stack (with Andy Pernsteiner—VAST Data)

    In episode 80, Andy Pernsteiner (field CTO VAST Data) joins us to discuss how VAST has evolved from a storage layer to application backend services, data management and AI infrastructure.The VAST Cosmos Community:https://community.vastdata.com/

  • S1 · E79
    June 26 · 30 min

    Databricks Summit 2026

    Tony reports from Databricks Data + AI Summit 2026. Lakehouse/RT, LTAP, Genie One, Omnigent, Unity AI Gateway, and more.

  • S1 · E78
    June 23 · 16 min

    What's New in Apache Spark?

    Lisa Cao (Databricks) joins us to discuss major new features in Apache Spark 4.1, and the continuing evolution of the platform to meet new challenges and serve an expanding user base.

  • S1 · E76
    June 23 · 17 min

    Database Development in The Era of Vibe Coding (with Nikita Shamgunov—Databricks)

    Nikita Shamgunov (Databricks) argues that all stateful systems in the future will need to support forking and version control to allow agents to make mistakes. Also on the agenda: wrangling spaghetti data, managing schema drift, and how the Lakebase architecture can outperform dedicated hardware. #artificialintelligence #aiagents #database

  • S1 · E75
    June 23 · 25 min

    Snowflake Summit 26 (with Juan Sequeda)

    Juan Sequeda joins Tony to go over the latest announcements and their broader implications for the industry. CoCo, CoWork, Horizon Context, data sharing, context wars, and more.#artificialintelligence #aiagents

Showing 1–20 of 20 episodes