Embedded AI Engineer, On-Device Models

Deepgram · Remote

Posted
47 days ago
Last confirmed live
1 day ago

What this role involves

Deepgram is hiring an Embedded AI Engineer to optimize and deploy speech models on resource-constrained devices. The role involves writing performance-critical code in C/C++/Rust, applying model compression techniques, and working with embedded platforms. The ideal candidate has experience with embedded systems and model optimization.

Skills this posting asks for

  • c
  • c++
  • rust
  • quantization
  • pruning
  • distillation
  • operator fusion
  • embedded socs
  • microcontrollers
  • ai accelerators
  • real-time operating systems
  • freertos
  • model optimization
  • on-device inference
  • performance-critical code
  • speech-to-text
  • text-to-speech
  • deep learning
  • compiler optimization
  • mobile application processors

Requirements

  • Level: senior

From the employer’s posting

COMPANY OVERVIEW Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,300+ organizations build vo…

Read the full description on Deepgram’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at Deepgram

All 21 roles at Deepgram