Description
Ema is hiring a Site Reliability Engineer to own the stability, availability, and operational health of its agentic AI platform across customer environments. The role involves provisioning cloud infrastructure, executing on-call SaaS deployments, monitoring and responding to production incidents, improving observability, and maintaining deployment documentation. Candidates should have 3–5 years of experience in DevOps, infrastructure, or deployment engineering, along with hands-on cloud, infrastructure-as-code, CI/CD, and observability experience.
