Listen Top Shows Blog

Cutting-Edge Data Engineering at Teya with Alexandre Magno Lima Martins

Cutting-Edge Data Engineering at Teya with Alexandre Magno Lima Martins

Update: 2024-08-08

Share

Description

Data engineering is constantly evolving and staying ahead means mastering tools like Apache Airflow. In this episode, we explore the world of data engineering with Alexandre Magno Lima Martins, Senior Data Engineer at Teya. Alexandre talks about optimizing data workflows and the smart solutions they've created at Teya to make data processing easier and more efficient.

Key Takeaways:

(02:01 ) Alexandre explains his role at Teya and the responsibilities of a data platform engineer.
(02:40 ) The primary use cases of Airflow at Teya, especially with dbt and machine learning projects.
(04:14 ) How Teya creates self-service DAGs for dbt models.
(05:58 ) Automating DAG creation with CI/CD pipelines.
(09:04 ) Switching to a multi-file method for better Airflow performance.
(12:48 ) Challenges faced with Kubernetes Executor vs. Celery Executor.
(16:13 ) Using Celery Executor to handle fast tasks efficiently.
(17:02 ) Implementing KEDA autoscaler for better scaling of Celery workers.
(19:05 ) Reasons for not using Cosmos for DAG generation and cross-DAG dependencies.
(21:16 ) Alexandre's wish list for future Airflow features, focusing on multi-tenancy.

Resources Mentioned:

Alexandre Magno Lima Martins -
https://www.linkedin.com/in/alex-magno/
Teya -
https://www.linkedin.com/company/teya-global/
Apache Airflow -
https://airflow.apache.org/
dbt -
https://www.getdbt.com/
Kubernetes -
https://kubernetes.io/
KEDA -
https://keda.sh/

Thanks for listening to The Data Flowcast: Mastering Airflow for Data Engineering & AI. If you enjoyed this episode, please leave a 5-star review to help get the word out about the show. And be sure to subscribe so you never miss any of the insightful conversations.

#AI #Automation #Airflow #MachineLearning

Comments

Top Podcasts

The Best New Comedy Podcast Right Now – June 2024 The Best News Podcast Right Now – June 2024 The Best New Business Podcast Right Now – June 2024 The Best New Sports Podcast Right Now – June 2024 The Best New True Crime Podcast Right Now – June 2024 The Best New Joe Rogan Experience Podcast Right Now – June 20 The Best New Dan Bongino Show Podcast Right Now – June 20 The Best New Mark Levin Podcast – June 2024

In Channel

Exploring the Power of Airflow 3 at Astronomer with Amogh Desai

Exploring the Power of Airflow 3 at Astronomer with Amogh Desai

2024-12-2030:24

Using Airflow To Power Machine Learning Pipelines at Optimove with Vasyl Vasyuta

Using Airflow To Power Machine Learning Pipelines at Optimove with Vasyl Vasyuta

2024-12-1224:11

Maximizing Business Impact Through Data at GlossGenius with Katie Bauer

Maximizing Business Impact Through Data at GlossGenius with Katie Bauer

2024-12-0525:49

Optimizing Large-Scale Deployments at LinkedIn with Rahul Gade

Optimizing Large-Scale Deployments at LinkedIn with Rahul Gade

2024-12-0227:47

How Uber Manages 1 Million Daily Tasks Using Airflow, with Shobhit Shah and Sumit Maheshwari

How Uber Manages 1 Million Daily Tasks Using Airflow, with Shobhit Shah and Sumit Maheshwari

2024-11-1428:44

Building Resilient Data Systems for Modern Enterprises at Astrafy with Andrea Bombino

Building Resilient Data Systems for Modern Enterprises at Astrafy with Andrea Bombino

2024-11-0728:29

Inside Airflow 3: Redefining Data Engineering with Vikram Koka

Inside Airflow 3: Redefining Data Engineering with Vikram Koka

2024-10-3130:08

Building a Data-Driven HR Platform at 15Five with Guy Dassa

Building a Data-Driven HR Platform at 15Five with Guy Dassa

2024-10-2420:25

The Intersection of AI and Data Management at Dosu with Devin Stein

The Intersection of AI and Data Management at Dosu with Devin Stein

2024-10-0420:18

AI-Powered Vehicle Automation at Ford Motor Company with Serjesh Sharma

AI-Powered Vehicle Automation at Ford Motor Company with Serjesh Sharma

2024-09-1226:11

From Task Failures to Operational Excellence at GumGum with Brendan Frick

From Task Failures to Operational Excellence at GumGum with Brendan Frick

2024-09-0624:06

From Sensors to Datasets: Enhancing Airflow at Astronomer with Maggie Stark and Marion Azoulai

From Sensors to Datasets: Enhancing Airflow at Astronomer with Maggie Stark and Marion Azoulai

2024-08-2922:25

Mastering Data Orchestration with Airflow at M Science with Ben Tallman

Mastering Data Orchestration with Airflow at M Science with Ben Tallman

2024-08-2624:36

Welcome to The Data Flowcast

Welcome to The Data Flowcast

2024-08-1902:01

Enhancing Business Metrics With Airflow at Artlist with Hannan Kravitz

Enhancing Business Metrics With Airflow at Artlist with Hannan Kravitz

2024-08-1523:51

Cutting-Edge Data Engineering at Teya with Alexandre Magno Lima Martins

Cutting-Edge Data Engineering at Teya with Alexandre Magno Lima Martins

2024-08-0823:46

Airflow Strategies for Business Efficiency at Campbell with Larry Komenda

Airflow Strategies for Business Efficiency at Campbell with Larry Komenda

2024-07-2626:10

How Laurel Uses Airflow To Enhance Machine Learning Pipelines with Vincent La and Jim Howard

How Laurel Uses Airflow To Enhance Machine Learning Pipelines with Vincent La and Jim Howard

2024-07-1823:58

How Vibrant Planet's Self-Healing Pipelines Revolutionize Data Processing

How Vibrant Planet's Self-Healing Pipelines Revolutionize Data Processing

2024-06-2723:51

The Future of AI in Data Engineering With Astronomer’s Julian LaNeve and David Xue

The Future of AI in Data Engineering With Astronomer’s Julian LaNeve and David Xue

2024-05-2923:36

00:00

00:00

x

Cutting-Edge Data Engineering at Teya with Alexandre Magno Lima Martins

Cutting-Edge Data Engineering at Teya with Alexandre Magno Lima Martins

support@contentallies.com (Astronomer)