Description
Newton is hiring a Site Reliability Engineer to improve the reliability, resilience, scalability, and operational readiness of its services. The role involves improving infrastructure, managing incidents, supporting critical services through on-call rotations, defining reliability metrics and error budgets, improving observability, debugging production systems, and leading postmortems. Candidates should have experience with scalable systems in AWS or a similar cloud environment, critical-system on-call experience, chaos engineering, live production debugging, scripting or development, and self-directed work in a growing environment.
