Description
Oracle is hiring a Distributed Systems Engineer to design, develop, and operate large-scale GPU infrastructure supporting AI/ML/HPC workloads. The role focuses on distributed system architecture, scalability, reliability, fault tolerance, data synchronization, performance and load testing, operational troubleshooting, incident response, security, and infrastructure automation. Candidates need a BS or MS in Computer Science or a relevant technical field, at least four years of experience delivering and operating large-scale production systems, and proficiency in one programming language. The role offers a US hiring range of $79,200 to $209,500 per annum, with bonus and equity eligibility and comprehensive benefits.
