Description
The Site Reliability Engineer supports Oracle Analytics Cloud services by improving reliability, availability, performance, operational support, and maturity. The role performs production and pre-production SRE activities, provides 24x7 follow-the-sun support, responds to incidents, troubleshoots complex distributed-system issues, investigates application and service code, develops automation and monitoring tooling, analyzes capacity and service health, executes patches and upgrades, and partners with development, support, product, and other engineering teams. The position requires a computer science or engineering degree or equivalent practical experience, cloud and distributed-systems experience, Linux/Unix administration, networking knowledge, and the ability to work independently in an agile environment.
