Skip to main content

GPU Systems Engineer – HPC / Parallel Computing at vastai

Department: Engineering

Compensation

$200,000 – $330,000/yr

Setup
On-site
Location
San Francisco, California
Type
Full-time
Level
not_specified
Posted

Description

Vast.ai is hiring a full-time Systems Engineer to scale AI inference by designing and optimizing GPU kernels and tensor libraries, translating HPC techniques into scalable inference solutions, evaluating emerging architectures and resource-management approaches, and improving GPU infrastructure efficiency. The role requires CUDA/C++, GPGPU, Python, and Linux, with advanced C++, parallel programming, systems optimization, and HPC performance tooling experience preferred. It is on-site in San Francisco or Westwood, Los Angeles, and includes health, dental, vision, life insurance, a 401(k) match, equity, and onsite meals.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation