Description
Join our SRE team as an enthusiastic and proactive Site Reliability Engineer to ensure world-class resilience and performance across the platform. This role involves advising on site reliability aspects like availability, scalability, observability, and capacity planning. You will proactively monitor platform performance, collaborate with engineering teams to resolve bottlenecks, implement SLOs, enhance observability, and champion best practices for high availability. Additionally, you'll conduct incident response, troubleshoot issues, participate in blameless post-mortems, and develop documentation.
