Description
tastytrade is hiring a Senior Site Reliability Engineer to build fault-tolerant, self-healing infrastructure and observability for its brokerage platform. The role focuses on improving telemetry, logging, alerting, SLOs, error budgets, and scalability across HashiCorp Nomad, while extending Prometheus, Honeycomb, and OpenTelemetry instrumentation. The engineer will work with Ruby, Java, and Elixir services, mentor other engineers, and support production on-call and post-incident processes. The position is hybrid in Chicago, Illinois, with three days per week in the office, and offers a base salary of $180,000-$200,000 plus a discretionary performance bonus.
