Staff Machine Learning Engineer - Foundation Model

xpengmotors · Santa Clara, CA

Posted
83 days ago
Last confirmed live
1 day ago

What this role involves

This role focuses on designing and implementing large-scale Vision-Language-Action (VLA) foundation models for end-to-end autonomous driving. The engineer will develop pretraining and fine-tuning strategies using multimodal fleet data, research cross-modal alignment techniques, and collaborate on scaling model training across thousands of GPUs.

Skills this posting asks for

  • machine learning
  • deep learning
  • transformers
  • vision-language models
  • multimodal learning
  • distributed training
  • fsdp
  • ddp
  • autonomous driving
  • python
  • pytorch
  • tensorflow
  • cuda
  • reinforcement learning
  • imitation learning
  • lane detection
  • object detection
  • trajectory prediction

Requirements

  • Level: staff

From the employer’s posting

<div class="ace-line ace-line old-record-id-TMEXdNB2hoez9lx0y…

Read the full description on xpengmotors’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at xpengmotors

All 14 roles at xpengmotors