Description
Verbit is hiring an ML/AI Engineer to build, deploy, and serve machine-learning models in production, with a focus on speech recognition and legal insights. The role includes optimizing inference performance, latency, and GPU utilization; supporting ASR integrations and self-hosted models; evaluating ASR and ML metrics; and optionally integrating LLMs and agentic flows. Candidates need at least four years of machine or deep learning experience, strong Python and Linux skills, familiarity with PyTorch or TensorFlow, MLOps and production-serving experience, GPU optimization, and experience with Docker, gRPC, serverless or microservices architectures, and real-time streaming systems.
