Platform Engineer - AI/ML Infrastructure (Kubernetes & Terraform)

Deepgram · Remote

Posted
51 days ago
Last confirmed live
1 day ago

What this role involves

Deepgram is seeking a Site Reliability Engineer to build and operate hybrid infrastructure for AI/ML workloads using Kubernetes, AWS, and Terraform. The role involves managing on-premise GPU servers, implementing observability and automation, and collaborating with research teams to accelerate model development. Candidates must be comfortable with rapid change and active use of AI tools.

Skills this posting asks for

  • kubernetes
  • aws
  • terraform
  • slurm
  • infrastructure-as-code
  • networking
  • storage
  • observability
  • gpu
  • ai/ml
  • automation
  • incident-response
  • cncf

Requirements

  • Level: senior

From the employer’s posting

COMPANY OVERVIEW Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,300+ organizations build vo…

Read the full description on Deepgram’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at Deepgram

All 21 roles at Deepgram