Description
Trainline is hiring an Embedded Data Engineer in Machine Learning to build scalable data pipelines, models, and feature stores for analytics and machine learning workloads. The role involves deploying and maintaining cloud-native applications on AWS, using CI/CD, observability, and engineering best practices, while collaborating with Machine Learning Engineers and Data Scientists. Required skills include Python, SQL, data pipelines, feature engineering, cloud data modelling, Spark, Airflow, and real-time or batch data workloads; experience with Ray, Parquet, Iceberg, Terraform, Docker, and CI/CD is preferred. The position follows a hybrid model requiring at least 60% office work over a 12-week period and includes private healthcare and dental insurance.
