Embedded AI Engineer, On-Device Models
Deepgram · Remote
- Posted
- 47 days ago
- Last confirmed live
- 1 day ago
What this role involves
Deepgram is hiring an Embedded AI Engineer to optimize and deploy speech models on resource-constrained devices. The role involves writing performance-critical code in C/C++/Rust, applying model compression techniques, and working with embedded platforms. The ideal candidate has experience with embedded systems and model optimization.
Skills this posting asks for
- c
- c++
- rust
- quantization
- pruning
- distillation
- operator fusion
- embedded socs
- microcontrollers
- ai accelerators
- real-time operating systems
- freertos
- model optimization
- on-device inference
- performance-critical code
- speech-to-text
- text-to-speech
- deep learning
- compiler optimization
- mobile application processors
Requirements
- Level: senior
From the employer’s posting
COMPANY OVERVIEW Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,300+ organizations build vo…
Read the full description on Deepgram’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at Deepgram
- Senior Data Scientist, Data FlywheelRemote
- Full Stack Web Developer, MarketingRemote
- Staff Product Manager, Agentic Experiences (Former Engineer)Remote
- Senior Software Engineer - Model Evaluation & AI SystemsRemote
- Staff Product Manager (Product-Led Growth)Remote
- Senior Product Manager, EnterpriseRemote