Description
The role is responsible for designing, implementing, and maintaining scalable and secure cloud infrastructure for AI/ML solutions, including Infrastructure as Code, CI/CD, MLOps, deployment automation, observability, performance optimization, production troubleshooting, and DevSecOps. It requires 5+ years of DevOps, SRE, or platform engineering experience, strong MLOps and cloud experience, and proficiency with tools such as Terraform, Kubernetes, Python, and monitoring platforms. The work environment is agile and collaborative, with in-person collaboration and flexibility to support individual work styles.
