Staff LLM Quantization & Deployment Engineer
XPENG in Santa Clara, CA, is seeking a Staff Machine Learning Engineer to advance LLM quantization and on-vehicle deployment. You will develop PTQ, QAT, mixed-precision inference, and robust production pipelines.
The role requires a Master in CS/EE, 3–5 years of experience, strong PyTorch and Python skills, and the ability to collaborate across research, systems, and product teams. We offer competitive compensation and ample resources.
#J-18808-Ljbffr