Description
The role leads a software and systems engineering team responsible for scaling alerting infrastructure, improving service reliability, automation, and self-repair for Google Cloud Memorystore. Responsibilities include team management, OKR execution, cross-functional partnerships, technical direction, oncall participation, and postmortem culture. The position requires a bachelor’s degree in computer science or a related field or equivalent practical experience, plus eight years of software development experience; distributed-systems, programming, and large-scale engineering experience are preferred.
