AI-powered
podcast player
Listen to all your favourite podcasts with AI-powered features
The Differences Between Open Source Data Diff and Cloud Product
Open source data diff is a utility tool that helps compare data across any database or within database at scale. It can be installed into the DBT project as a standalone package, but it reuses a lot of the configuration and credentials from the local DBT setup. Now the cloud product is focused on enabling teams. And so when it comes to automating testing for teams, it's really important to one, make sure that the testing is not necessarily done by individual engineers,. But it's actually performed by an automated process as part of continuous integration.
Data engineering is all about building workflows, pipelines, systems, and interfaces to provide stable and reliable data. Your data can be stable and wrong, but then it isn't reliable. Confidence in your data is achieved through constant validation and testing. Datafold has invested a lot of time into integrating with the workflow of dbt projects to add early verification that the changes you are making are correct. In this episode Gleb Mezhanskiy shares some valuable advice and insights into how you can build reliable and well-tested data assets with dbt and data-diff.
The intro and outro music is from The Hug by The Freak Fandango Orchestra / CC BY-SA
Special Guest: Gleb Mezhanskiy.
Sponsored By:
Listen to all your favourite podcasts with AI-powered features
Listen to the best highlights from the podcasts you love and dive into the full episode
Hear something you like? Tap your headphones to save it with AI-generated key takeaways
Send highlights to Twitter, WhatsApp or export them to Notion, Readwise & more
Listen to all your favourite podcasts with AI-powered features
Listen to the best highlights from the podcasts you love and dive into the full episode