Platform Engineer - AI/ML Infrastructure (Kubernetes & Terraform)
Deepgram · Remote
- Posted
- 51 days ago
- Last confirmed live
- 1 day ago
What this role involves
Deepgram is seeking a Site Reliability Engineer to build and operate hybrid infrastructure for AI/ML workloads using Kubernetes, AWS, and Terraform. The role involves managing on-premise GPU servers, implementing observability and automation, and collaborating with research teams to accelerate model development. Candidates must be comfortable with rapid change and active use of AI tools.
Skills this posting asks for
- kubernetes
- aws
- terraform
- slurm
- infrastructure-as-code
- networking
- storage
- observability
- gpu
- ai/ml
- automation
- incident-response
- cncf
Requirements
- Level: senior
From the employer’s posting
COMPANY OVERVIEW Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,300+ organizations build vo…
Read the full description on Deepgram’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at Deepgram
- Senior Data Scientist, Data FlywheelRemote
- Full Stack Web Developer, MarketingRemote
- Staff Product Manager, Agentic Experiences (Former Engineer)Remote
- Senior Software Engineer - Model Evaluation & AI SystemsRemote
- Staff Product Manager (Product-Led Growth)Remote
- Senior Product Manager, EnterpriseRemote