Description
interface.ai is hiring a senior-most, hands-on IC Site Reliability Engineer to own the reliability, security, and operational standards of its production AI platform serving banks and credit unions. The role covers SLOs and error budgets, disaster recovery, GitOps delivery, AWS and Kubernetes infrastructure, incident management, observability, AI-native automation, and infrastructure security for SOC 2, FFIEC, and other regulated reviews. Candidates need production systems experience, deep AWS Kubernetes and GitOps expertise, infrastructure-as-code, SLO and DR practices, security controls, strong TypeScript or Python and Bash, and a BS/BA in Computer Science; the role is based onsite in San Francisco with on-call participation.
