Summary from listing
The role focuses on building and operating scalable data pipelines using Python, Apache Spark, AWS Glue ETL, and data-lake concepts. It involves extracting data from legacy and modern sources, engineering high-performance pipelines, and providing expert-level production support. The position also requires infrastructure and CI/CD experience with CloudFormation, Terraform, and GitHub Actions.

