Description
Nace AI is hiring a Senior MLOps Engineer in Palo Alto, California, for a full-time on-site role. The position owns infrastructure that takes small language models from research to production, including training orchestration, experiment tracking, model registries, CI/CD, automated evaluation, LLM/SLM serving, GPU cluster management, observability, inference-time optimization, and enterprise auditability. The role requires at least five years of MLOps, ML infrastructure, or platform engineering experience, production LLM serving experience, strong Kubernetes and Python skills, GPU cluster management, and a BS in computer science or a related technical field.
