Skip to main content

LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer at Qualcomm International Inc., Mexico Branch Office

Department: Machine Learning Engineering

Language

Compensation

$162,000 – $243,000/yr

Setup
On-site
Location
San Diego, California · Markham, Ontario
Type
Full-time
Level
not_specified
Posted

Description

Qualcomm is hiring LLM Serving Engineers to build scalable LLM inference platforms and contribute to serving packages such as vLLM, SGLang, Triton-Inference Server, and others. The role covers inference techniques, model optimization, autoscaling, load balancing, routing, customer collaboration, and open-source contributions. Candidates need strong Python, PyTorch, distributed systems, computer architecture, and deep-learning optimization experience, along with a relevant bachelor's, master's, or PhD degree and the specified years of related work experience. The listed pay range is $162,000 to $243,000.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation