Skip to main content

Applied ML Engineer - Edge Devices at Deepgram

Department: Engineering

Compensation

$155,000 – $245,000/yr

Setup
Remote
Type
Full-time
Level
senior
Posted

Description

Deepgram is hiring an Applied ML Engineer on its Partner Platform Engineering team to port speech models to non-NVIDIA and edge platforms. The role adapts model structure, operators, quantization, precision, and architecture to fit target hardware, validates accuracy and latency, builds repeatable deployment pipelines, and collaborates with embedded AI engineers and silicon vendors. It requires production experience deploying ML models to edge or non-NVIDIA hardware, knowledge of quantization and inference runtimes, Python and PyTorch, and the ability to automate model conversion and deployment.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation