Description
The Site Reliability Engineer will improve the reliability, observability, performance, and availability of critical online-gambling systems. The role involves incident resolution, service instrumentation with OpenTelemetry, logging, automation, infrastructure-as-code, observability dashboards, AI-assisted operations, and collaboration across software development and IT operations. It requires Python, Golang, JavaScript, SRE, observability, shell scripting, Ansible, Terraform, and experience with large-scale 24/7 enterprise systems, and the position is eligible for hybrid working from home.
