Skip to main content

Research Engineer, Infrastructure, Training Systems at ThinkingMachines

Department: Research Infrastructure (ML Infrastructure and Training Stack)

Compensation

$350,000 – $475,000/yr

Setup
On-site
Location
San Francisco, California
Type
Full-time
Level
not_specified
Posted

Description

Thinking Machines is hiring an Infrastructure Research Engineer to design, implement, and optimize distributed training systems for large-scale model training across thousands of GPUs and nodes. The role focuses on high-performance training, reusable frameworks, reliability, maintainability, security, and collaboration with researchers and engineers, with an emphasis on making experimentation and training fast and reliable. The position is based in San Francisco, California, offers visa sponsorship, and provides health, dental, vision, PTO, parental leave, and relocation benefits.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation