Staff Software Engineer - GenAI inference

Databricks · San Francisco, California · $191k–$233k

Posted
5 days ago
Last confirmed live
Today
Published range
$191k–$233k

What this role involves

Lead the architecture and optimization of the GenAI inference engine at Databricks. Responsible for kernel-level performance, distributed systems integration, and bridging research with production. Requires deep expertise in GPU programming, ML inference, and large-scale system design.

Skills this posting asks for

  • cuda
  • gpu programming
  • cublas
  • cudnn
  • nccl
  • rpc
  • distributed systems
  • ml inference
  • quantization
  • profiling
  • tracing
  • open source contributions
  • model serving
  • attention
  • mlp
  • recurrent modules
  • sparse operations
  • memory management
  • scheduling
  • batching
  • orchestration

Requirements

  • 6 years of experience
  • Level: staff

From the employer’s posting

P-1285 About This Role As a staff software engineer for GenAI inference, you will lead the architecture, development, and optimization of the inference engine that powers Databricks Foundation Model API.. You’ll bridge research advances and production demands, en…

Read the full description on Databricks’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at Databricks

All 216 roles at Databricks