Description
Propio is hiring a Senior Machine Learning Engineer, Speech & LLM Training Data to own the data roadmap and productionize secure, reproducible workflows for multilingual speech, translation, multimodal LLMs, and conversational AI. The role covers audio-processing pipelines, dataset curation and quality workflows, annotation guidelines, model training and evaluation, and experimentation on AWS. Candidates need a relevant bachelor’s or master’s degree, at least five years of experience in ML engineering, speech/audio ML, data engineering, NLP, or LLM training-data workflows, and strong experience with Python, SQL, Linux, Git, Docker, PyTorch, Hugging Face, FFmpeg, speech-processing tools, Databricks/Spark, AWS, annotation platforms, and experiment-tracking tools.

