Description
Oracle is hiring an AI Infrastructure Engineering Manager to lead teams building and scaling distributed AI compute infrastructure, including GPU control and data planes. The role oversees system scalability, reliability, fault tolerance, SLOs, observability, incident management, security, compliance, automation, and change management while setting team goals, coaching engineers, and tracking OKRs. Candidates need at least 10 years of software development experience, 5 years in people management, and 5 years designing large-scale distributed systems, with a relevant degree or equivalent experience. The role offers a US hiring range of $146,300 to $306,400 per annum plus potential bonus, equity, and compensation deferral.
