Description
Qualcomm is hiring AI Performance Engineers for its Cloud AI Engineering team to optimize and deploy language, vision, and diffusion models for efficient inference. The role involves PyTorch and ONNX model optimization, GenAI algorithm analysis, throughput and latency scaling, hardware workload mapping, root-cause analysis, and kernel development in Triton. Candidates need strong Python, transformer and inference-optimization knowledge, relevant hardware or systems experience, and a bachelor's, master's, or PhD in a related field. The listed pay range is $194,400 to $291,600.
