Description
The role is a customer-facing Site Reliability Engineer working with complex cloud platforms such as Kubernetes, KubeVirt, Cilium, Ceph, and Talos. Responsibilities include analyzing and improving workload reliability, defining SLOs, debugging complex issues with customers, automating recurring problems, and supporting security and compliance in regulated environments. The position requires a computer science degree or comparable qualification, at least two years of cloud infrastructure experience with Kubernetes, software development knowledge, and willingness to participate in an on-call schedule.
