Staff Software Engineer: AI Inference Data Plane

digitalocean98 · Seattle · $167k–$209k

Posted
4 days ago
Last confirmed live
Today
Published range
$167k–$209k

What this role involves

DigitalOcean seeks a Staff Engineer to design and optimize serverless AI inference infrastructure and APIs. The role involves building scalable multi-tenant services, improving observability and reliability, and collaborating with platform and product teams. Candidates need 8+ years in distributed systems, strong Go and Kubernetes experience, and familiarity with LLM serving architectures.

Skills this posting asks for

  • go
  • golang
  • kubernetes
  • distributed systems
  • microservices
  • cloud-native
  • sre
  • observability
  • incident management
  • reliability engineering
  • capacity planning
  • operational automation
  • gpu utilization
  • ttft
  • tpot
  • vllm
  • triton
  • api gateways
  • service mesh
  • tensorrt-llm
  • inference optimization
  • rate limiting
  • workload orchestration

Requirements

  • 8 years of experience
  • Level: staff
  • Remote policy: hybrid

From the employer’s posting

Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized…

Read the full description on digitalocean98’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at digitalocean98

All 29 roles at digitalocean98