Summary from listing
The employer is seeking a Data Engineer with at least six years of experience, including hands-on work building and operating large-scale online services, big-data pipelines, or middleware platforms using public cloud, distributed clusters, microservices, and cloud-native technologies. The role requires Hadoop and Spark, Python, Linux and shell scripting, containerization with Docker and Kubernetes, workflow engines such as Apache Argo, cloud platforms including GCP, AWS, or Azure, and experience with EMR, RDS, Redshift, model lifecycle management, machine learning frameworks, and relational and NoSQL databases.
