Staff Software Engineer, Infrastructure Engineering

coreweave · New York, NY / Sunnyvale, CA

Posted
2 days ago
Last confirmed live
Today

What this role involves

The role is within the Compute Architecture organization at CoreWeave, focusing on building Go-based distributed services for large-scale GPU data center infrastructure. Responsibilities include designing automation for hardware lifecycle management, improving observability, and ensuring reliability at fleet scale. The position requires strong Go expertise, experience with Kubernetes and observability stacks, and a track record of technical leadership.

Skills this posting asks for

  • go
  • rest
  • grpc
  • kubernetes
  • cloud-native
  • distributed systems
  • prometheus
  • grafana
  • promql
  • ci/cd
  • gpu servers
  • bmc
  • firmware
  • api development
  • observability
  • alerting
  • postmortems
  • incident response
  • mentoring
  • technical leadership

Requirements

  • 8 years of experience
  • Level: staff

From the employer’s posting

CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of…

Read the full description on coreweave’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at coreweave

All 88 roles at coreweave