Staff/Principal DevOps Engineer, AI Inference

lilasciences · Cambridge, MA USA

Posted
26 days ago
Last confirmed live
1 day ago

What this role involves

This role focuses on building and optimizing infrastructure for serving machine learning models at scale, including GPU clusters and cloud accelerators. The engineer will work on Kubernetes-based inference platforms, model serving frameworks, and observability systems. Collaboration with ML engineers and research scientists is key.

Skills this posting asks for

  • kubernetes
  • gpu
  • aws
  • terraform
  • helm
  • python
  • vllm
  • triton inference server
  • tgi
  • nvidia
  • aws inferentia
  • aws trainium
  • docker
  • cuda
  • nccl
  • rust
  • go

Requirements

  • Level: staff

From the employer’s posting

Your Impact at LILA The Staff/Principal DevOps Engineer - AI Inference will drive the design, implementation, and optimization of infrastructure purpose-built for serving machine learning models at scale. This role bridges platform engineering, site reliability, and ML in…

Read the full description on lilasciences’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at lilasciences

All 34 roles at lilasciences