Data Engineering Podcast


This show goes behind the scenes for the tools, techniques, and difficulties associated with the discipline of data engineering. Databases, workflows, automation, and data manipulation are just some of the topics that you will find here.

Support the show!

Rewind 10 seconds
1X
Skip 30 seconds ahead
0:00/0:00

Listen in your favorite app:



More options

Here are shows you might like

See show recommendations
AI Engineering Podcast
Tobias Macey
The Python Podcast.__init__
Tobias Macey

487 Episodes

Building Your Data Warehouse On Top Of PostgreSQL - E186

Summary

There is a lot of attention on the database market and cloud data warehouses. While they provide a measure of convenience, they also require you to sacrifice a certain amount of control over your data. If you want to build a warehouse that gives you both control and flexibility then you…

Summary

There is a lot of attention on the…

14 May 2021 | 01:15:07


Making Analytical APIs Fast With Tinybird - E185

Summary

Building an API for real-time data is a challenging project. Making it robust, scalable, and fast is a full time job. The team at Tinybird wants to make it easy to turn a continuous stream of data into a production ready API or data product. In this episode CEO Jorge Sancha explains how…

Summary

Building an API for real-time data is a…

11 May 2021 | 00:54:24


Making Spark Cloud Native At Data Mechanics - E184

Summary

Spark is one of the most well-known frameworks for data processing, whether for batch or streaming, ETL or ML, and at any scale. Because of its popularity it has been deployed on every kind of platform you can think of. In this episode Jean-Yves Stephan shares the work that he is doing at…

Summary

Spark is one of the most well-known…

07 May 2021 | 00:40:16


The Grand Vision And Present Reality of DataOps - E183

Summary

The Data industry is changing rapidly, and one of the most active areas of growth is automation of data workflows. Taking cues from the DevOps movement of the past decade data professionals are orienting around the concept of DataOps. More than just a collection of tools, there are a number…

Summary

The Data industry is changing rapidly,…

04 May 2021 | 00:57:08


Self Service Data Exploration And Dashboarding With Superset - E182

Summary

The reason for collecting, cleaning, and organizing data is to make it usable by the organization. One of the most common and widely used methods of access is through a business intelligence dashboard. Superset is an open source option that has…

27 April 2021 | 00:47:25


Moving Machine Learning Into The Data Pipeline at Cherre - E181

Summary

Most of the time when you think about a data pipeline or ETL job what comes to mind is a purely mechanistic progression of functions that move data from point A to point B. Sometimes, however, one of those transformations is actually a full-fledged machine learning project in its own right.…

Summary

Most of the time when you think about a…

20 April 2021 | 00:48:05


Exploring The Expanding Landscape Of Data Professions with Josh Benamram of Databand - E180

Summary

"Business as usual" is changing, with more companies investing in data as a first class concern. As a result, the data team is growing and introducing more specialized roles. In this episode Josh Benamram, CEO and co-founder of Databand, describes the motivations for these…

Summary

"Business as usual" is…

13 April 2021 | 01:08:36


Put Your Whole Data Team On The Same Page With Atlan - E179

Summary

One of the biggest obstacles to success in delivering data products is cross-team collaboration. Part of the problem is the difference in the information that each role requires to do their job and where they expect to find it. This introduces a barrier to communication that is difficult to…

Summary

One of the biggest obstacles to success…

06 April 2021 | 00:57:37


Data Quality Management For The Whole Team With Soda Data - E178

Summary

Data quality is on the top of everyone’s mind recently, but getting it right is as challenging as ever. One of the contributing factors is the number of people who are involved in the process and the potential impact on the business if something goes wrong. In this episode Maarten…

Summary

Data quality is on the top of…

30 March 2021 | 00:58:00


Real World Change Data Capture At Datacoral - E177

Summary

The world of business is becoming increasingly dependent on information that is accurate up to the minute. For analytical systems, the only way to provide this reliably is by implementing change data capture (CDC). Unfortunately, this is a non-trivial undertaking, particularly for teams…

Summary

The world of business is becoming…

23 March 2021 | 00:49:58