Summary from listing
The role is a data-engineering position focused on end-to-end data pipelines, including ingestion, transformation, consumption, data lake and lakehouse architecture, API integration, and technical design documentation. It requires expertise in SQL, Spark, Python, Scala, Hadoop, AWS, EMR, and related big-data technologies, along with experience in data mining, databases, debugging, stakeholder collaboration, and solution delivery. The candidate should be able to work independently, communicate effectively, identify performance indicators, and lead technical aspects of projects.
