Staff Machine Learning Engineer - AI Foundation

xpengmotors · Santa Clara, CA

Posted
83 days ago
Last confirmed live
1 day ago

What this role involves

This role focuses on building and optimizing ML infrastructure for large foundation models, particularly for autonomous driving. Responsibilities include optimizing transformer-based LLMs for low-latency inference, kernel optimization, and deploying models across various hardware. The candidate should have 5-8 years of industry experience and a Master's degree in CS/CE/EE.

Skills this posting asks for

  • transformer-based llms
  • cuda
  • triton
  • quantization
  • knowledge distillation
  • pruning
  • kv-cache optimization
  • gpus
  • cpus
  • edge accelerators

Requirements

  • 5 years of experience
  • Level: staff

From the employer’s posting

<div class="ace-line ace-line old-record-id-TMEXdNB2hoez9lx0y…

Read the full description on xpengmotors’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at xpengmotors

All 14 roles at xpengmotors