Description
Deepgram is hiring a Research Staff Data Scientist to build scalable data pipelines and systems for speech and language AI foundation models. The role focuses on acquiring, preparing, and synthesizing conversational audio data; characterizing complex audio with signal-processing and deep-learning methods; automating human annotation and feedback systems; developing benchmarks and curated datasets; and presenting experimental results. Candidates should have experience with data processing pipelines, statistical methods, deep learning, Python, and PyTorch, with speech, audio, or language-processing backgrounds preferred.
