Description
Vapi is hiring a Senior Staff Site Reliability Engineer to join its five-person Infrastructure team in San Francisco. The role focuses on reliability and operability across compute, storage, networking, and telephony systems, with responsibilities including building reliability tooling and automation, improving observability, incident response, capacity, performance, and production automation, and delivering durable improvements to failure prevention and recovery. The engineer should be a senior or staff-level software engineer with distributed-systems experience, production-quality software development skills, and expertise in observability, incident response, failure analysis, Kubernetes, networking, and cloud infrastructure.
