Description
OpenAI is hiring for a general Compute Infrastructure role to build and operate reliable, scalable systems supporting frontier AI research and products. The work spans heterogeneous hardware, networking, storage, cluster orchestration, scheduling, fleet health, CaaS, agent infrastructure, observability, benchmarking, and workload optimization. Candidates should have strong software engineering experience with production infrastructure systems and relevant expertise in distributed systems, networking, Kubernetes, GPU infrastructure, reliability, or related areas.
