Senior/Staff AI Engineer

DDN · Remote

Posted
22 days ago
Last confirmed live
1 day ago

What this role involves

This role focuses on building and optimizing LLM serving and inference systems for production environments, with an emphasis on performance across GPU and CPU pathways, KV cache, memory, storage, and throughput bottlenecks. The ideal candidate has deep hands-on experience in production AI systems, particularly at the systems layer, and is comfortable working on infrastructure where storage architecture and systems efficiency materially affect AI performance.

Skills this posting asks for

  • llm serving
  • inference systems
  • gpu
  • cpu
  • kv cache
  • memory
  • storage
  • rag
  • retrieval
  • distributed systems
  • model serving
  • caching
  • distributed performance
  • ai infrastructure
  • high-performance systems
  • storage platforms

Requirements

  • Level: senior

From the employer’s posting

WHAT YOU’LL DO - Build and optimize LLM serving and inference systems for production environments - Improve performance across GPU and CPU pathways - Work on KV cache, memory, storage, and throughput bottlenecks - Design and scale systems that support RAG and retrieval-heavy AI workloads…

Read the full description on DDN’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at DDN

All 13 roles at DDN