Skip to main content

Site Reliability Engineer / Production Engineer, SLO and Observability at cloudlinux-1

Setup
Remote
Location
Pecinci Municipality, Vojvodina
Type
Full-time
Level
senior
Posted

Description

CloudLinux is hiring a remote-first SRE/production-engineering specialist to establish and operate an SLO, SLI, telemetry collection, alerting, and escalation framework for the Imunify360 Linux security product. The role covers defining measurable health indicators for approximately 70 cloud and fleet components, building telemetry pipelines and instrumentation in Python, Go, and Rust, implementing SLO-based burn-rate alerting and ownership routing, and coaching engineering squads on on-call and incident practices. The position is highly collaborative and focused on detecting silent security-control degradation within hours rather than months.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation