Description
Spotify is hiring a Senior Applied Research Scientist for its Speak team to develop novel machine-learning techniques and architectures for generative conversational speech-to-speech models. The role involves research in speech synthesis and recognition, end-to-end model development, production pipeline collaboration, infrastructure and data improvements, and scaling models for Spotify’s platform. Candidates should have a PhD and professional experience in machine learning, including work with transformers, GANs, diffusion models, flow matching, VAEs, or audio codecs, along with strong Python and PyTorch skills. The position is based in London or Stockholm and offers flexibility to work from home with some in-person meetings.
