Skip to main content

GPU Cluster Engineer, Systems & Platform at sciforium

Department: Engineering

Compensation

$150,000 – $220,000/yr

Setup
On-site
Location
San Francisco, California
Type
Full-time
Level
senior
Posted

Description

Sciforium is hiring a GPU Cluster Engineer to own the software stack of GPU clusters, from kernel and driver tuning through schedulers, containers, and ML frameworks. The role builds and validates production-ready node images, automates fleet upgrades and self-healing workflows, manages Kubernetes and Slurm/Run:AI workloads, maintains NVIDIA and AMD accelerator stacks, and troubleshoots distributed training and inference performance issues. Candidates need at least five years of systems or infrastructure engineering experience, a relevant bachelor's or master's degree, deep Linux and GPU-stack expertise, and experience with configuration management, provisioning, containers, and distributed filesystems.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation