Description
Plaid is hiring a Data Infrastructure Engineer to lead data platform projects involving data warehouses, data lakehouses, Apache Spark, streaming, workflow orchestration, and ETL pipelines. The role contributes to technical roadmaps, improves machine-learning development workflows, builds offline streaming and new ETL infrastructure, reduces operational burden, and mentors and leads junior engineers. Candidates need at least six years of software engineering experience, strong data infrastructure or platform experience, Python proficiency, and leadership experience; Databricks, Airflow, and AWS EMR experience are preferred.
