Staff/Principal DevOps Engineer, AI Inference
lilasciences · Cambridge, MA USA
- Posted
- 26 days ago
- Last confirmed live
- 1 day ago
What this role involves
This role focuses on building and optimizing infrastructure for serving machine learning models at scale, including GPU clusters and cloud accelerators. The engineer will work on Kubernetes-based inference platforms, model serving frameworks, and observability systems. Collaboration with ML engineers and research scientists is key.
Skills this posting asks for
- kubernetes
- gpu
- aws
- terraform
- helm
- python
- vllm
- triton inference server
- tgi
- nvidia
- aws inferentia
- aws trainium
- docker
- cuda
- nccl
- rust
- go
Requirements
- Level: staff
From the employer’s posting
Your Impact at LILA The Staff/Principal DevOps Engineer - AI Inference will drive the design, implementation, and optimization of infrastructure purpose-built for serving machine learning models at scale. This role bridges platform engineering, site reliability, and ML in…
Read the full description on lilasciences’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at lilasciences
- Senior Director, Data Platform EngineeringSan Francisco, CA USA
- Principal, Machine Learning EngineerSan Francisco, CA USA
- Sr Principal/ Principal Software Engineer, Scientific System of RecordCambridge, MA USA; San Francisco, CA USA
- Staff Software Engineer, Lab SoftwareCambridge, MA USA
- Software Engineer I, Instrument Software Cambridge, MA USA
- Senior Software Engineer, Operations ResearchCambridge, MA USA