Staff Software Engineer: AI Inference Data Plane
digitalocean98 · San Francisco · $167k–$209k
- Posted
- 4 days ago
- Last confirmed live
- Today
- Published range
- $167k–$209k
What this role involves
This role involves designing and optimizing serverless AI inference infrastructure and APIs at DigitalOcean. Responsibilities include building scalable multi-tenant services, improving platform resiliency, and providing technical leadership. Candidates need 8+ years of experience in distributed systems, proficiency in Go and Kubernetes, and deep SRE and observability knowledge.
Skills this posting asks for
- go
- kubernetes
- observability
- sre
- distributed systems
- microservices
- cloud-native
- vllm
- triton
- api gateways
- service mesh
- tensorrt-llm
- inference optimization
- rate limiting
- workload orchestration
- ttft
- tpot
- gpu utilization
- incident management
- capacity planning
- operational automation
Requirements
- 8 years of experience
- Level: senior
- Remote policy: remote
From the employer’s posting
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized…
Read the full description on digitalocean98’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at digitalocean98
- Software Engineer, AutomationSeattle
- Staff Product Manager, Compute Infrastructure and HardwareSeattle
- Principal Product Manager, Agentic AI InfrastructureSeattle
- Staff Product Manager, IAM/Identity ProductsSeattle
- Staff Product Manager, IAM/Identity ProductsBoston
- Staff Product Security Engineer, Secure Design (Kernel and Virtualization)United States