Skip to main content

Performance Engineer at Inferact

Compensation

$200,000 – $400,000/yr

Setup
Remote
Location
San Francisco, California
Type
Full-time
Posted

Description

Inferact is seeking a Performance Engineer to optimize vLLM, the world's AI inference engine, by developing and optimizing kernels for various accelerators like NVIDIA GPUs. The role involves working directly with hardware vendor teams to maximize performance across generations of hardware.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation