Skip to main content

GPU Kernel Engineer at FriendliAI

Department: Engineering

Setup
Hybrid
Location
San Francisco, California
Type
Full-time
Level
mid
Posted

Description

FriendliAI is hiring a GPU Kernel Engineer to design, implement, and optimize low-level compute kernels for its large-scale GPU-accelerated AI inference platform. The role focuses on high-performance CUDA and C++ kernel development, reduced-precision and quantized inference, cross-vendor NVIDIA and AMD performance tuning, GPU libraries, and multi-modal model pipelines. Candidates need at least three years of GPU programming or performance-critical systems experience, a relevant bachelor’s or master’s degree, strong CUDA or ROCm/HIP expertise, and deep knowledge of GPU architecture.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation