Skip to main content

Staff Machine Learning Engineer - LLM Quantization & Deployment at XPENG

Department: AI Infrastructure Team

Compensation

$215,280 – $364,320/yr

Setup
Hybrid
Location
Santa Clara, California
Type
Full-time
Level
not_specified
Posted

Description

XPENG is seeking a Lead Research Engineer specializing in Large Language Model (LLM) Deployment and Quality Sign-off. The role involves developing and optimizing VLA inference models, ensuring numerical consistency, and producing high-quality Python code for production environments. The engineer will collaborate with various internal teams to enhance model performance, curate evaluation datasets, analyze numerical errors, and contribute to advancements in autonomous driving.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation