Senior Machine Learning Engineer (Inference Platform)

wizardcommerce · Remote

Posted
84 days ago
Last confirmed live
2 days ago

What this role involves

This role involves owning the end-to-end lifecycle of production ML serving systems, focusing on the inference platform for a conversational shopping agent. The engineer will optimize performance, reliability, and scalability of multiple serving engines. Key responsibilities include building observability, enforcing SLAs, and collaborating with cross-functional teams.

Skills this posting asks for

  • python
  • aws
  • gcp
  • azure
  • vllm
  • tgi
  • tensorrt-llm
  • sglang
  • llm
  • continuous batching
  • kv-cache
  • quantization
  • ci/cd
  • gpu
  • ml lifecycle
  • model registries
  • inference serving
  • embedding models
  • extraction models
  • observability
  • monitoring
  • alerting
  • testing
  • docker

Requirements

  • 5 years of experience
  • Level: senior

From the employer’s posting

About Wizard AI At Wizard AI, we’re building the top-performing AI Shopping Agent that delivers the best products from across the web with unmatched accuracy, quality, and trust. Our ML models power the core of our platform, and we’re looking for a Senior M…

Read the full description on wizardcommerce’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at wizardcommerce

All 3 roles at wizardcommerce