Description
Arcesium is hiring a principal-level infrastructure and site reliability engineer to lead the architectural direction, reliability, scalability, and operational maturity of a Kubernetes-native, AWS-hosted front-office Portfolio and Order Management System. The hands-on individual contributor will own Kubernetes, ArgoCD, PostgreSQL, and AWS infrastructure; improve maintenance, CI/CD, Terraform workflows, monitoring, alerting, observability, and developer tooling; resolve complex production incidents; support HA/DR, security, networking, and standardization; and establish sustainable global on-call operations. The role requires a computer science or engineering degree and 9+ years of professional engineering experience, including significant principal-level or equivalent experience.
