Skip to main content

RESEARCHER, EFFICIENT INFERENCE at makermaker.ai

Department: Technical Staff

Setup
On-site
Location
San Francisco, California
Type
Full-time
Level
senior
Posted

Description

The company is hiring a senior research engineer to research and develop efficient machine-learning inference methods, including quantization, speculative decoding, distillation, sparse and structured attention, mixture-of-experts, and related training-time techniques. The role involves algorithm design, production-scale experimentation, evaluation, and close collaboration with inference engineering and model-research teams to move methods from prototypes into production. Candidates should have at least five years of hands-on research experience, strong training- and inference-performance knowledge, PyTorch or Jax expertise, statistical rigor, and published research at top machine-learning venues.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation