Staff Software Engineer - GenAI inference
Databricks · San Francisco, California · $191k–$233k
- Posted
- 5 days ago
- Last confirmed live
- Today
- Published range
- $191k–$233k
What this role involves
Lead the architecture and optimization of the GenAI inference engine at Databricks. Responsible for kernel-level performance, distributed systems integration, and bridging research with production. Requires deep expertise in GPU programming, ML inference, and large-scale system design.
Skills this posting asks for
- cuda
- gpu programming
- cublas
- cudnn
- nccl
- rpc
- distributed systems
- ml inference
- quantization
- profiling
- tracing
- open source contributions
- model serving
- attention
- mlp
- recurrent modules
- sparse operations
- memory management
- scheduling
- batching
- orchestration
Requirements
- 6 years of experience
- Level: staff
From the employer’s posting
P-1285 About This Role As a staff software engineer for GenAI inference, you will lead the architecture, development, and optimization of the inference engine that powers Databricks Foundation Model API.. You’ll bridge research advances and production demands, en…
Read the full description on Databricks’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at Databricks
- Specialist Solutions Architect - Cloud Infrastructure & Platform (AWS)United States
- Specialist Solutions Architect - Cloud Infrastructure & Platform (Azure)United States
- Software Engineering Intern (2027 Start) - WinterBellevue, Washington; Mountain View, California; San Francisco, California
- Solutions Architect - Communications, Media, Entertainment and Games Remote
- Solutions Architect - Communications, Media, Entertainment and GamesRemote
- Sr. Solutions Architect - Retail, Travel & HospitalityRemote