Skip to main content

Distributed Software Engineer at Cerebras Systems

Department: Software Engineering

Language
Setup
Hybrid
Location
Toronto, Ontario
Level
senior
Posted

Description

Cerebras Systems is hiring a Cluster Engineer to build and operate software that manages thousands of Cerebras wafers, servers, and switches as a cloud. The role covers declarative CRD-driven automation, bare-metal networking, OS, application software, cluster installation and security patching, Kubernetes operators for inference workloads, gRPC control-plane services, metrics and log pipelines, failure detection, high availability, and recovery. The position requires at least five years of production distributed systems or infrastructure software experience, strong Go and Python skills, deep Kubernetes knowledge, distributed-systems debugging ability, Prometheus and Grafana expertise, and demonstrated use of AI in engineering workflows.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation