Skip to main content

ML Engineer - Inference & Model Deployment at HiringCafe.com

Department: Founding Engineer

Compensation

$250,000 – $310,000/yr

Setup
On-site
Location
Cupertino, California
Type
Full-time
Level
staff+
Posted

Description

HiringCafe.com is hiring a founding ML engineer to build and operate production AI and machine learning systems. The role focuses on deploying researcher-trained models, profiling and optimizing inference latency, throughput, memory, and compute, implementing quantization and other optimization techniques, building scalable multi-GPU serving systems, and designing reliable architectures for millions of users. The position is based in Cupertino and requires in-person work, with health, dental, vision, parental leave, and relocation support.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation