Skip to main content

Research Engineer - AI Performance & Kernel Optimization at Zyphra

Department: R&D - Engineering

Setup
On-site
Location
San Francisco, California
Level
not_specified
Posted

Description

ZYPHRA is hiring a Research Engineer - AI Performance & Kernel Optimization to improve the performance of large-scale language model training and inference systems. The role focuses on developing and optimizing GPU kernels, tuning training and inference stacks, profiling bottlenecks, improving distributed training and inference, and optimizing performance across NVIDIA and non-NVIDIA accelerator platforms. Candidates should have strong systems and low-level performance intuition, experience with GPU kernel development and large-scale ML workloads, and familiarity with distributed training, parallelism, profiling, and hardware-software interactions.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation