Senior/Staff AI Engineer
DDN · Remote
- Posted
- 23 days ago
- Last confirmed live
- 1 day ago
What this role involves
This role focuses on building and optimizing LLM serving and inference systems for production environments, with an emphasis on performance across GPU and CPU pathways, KV cache, memory, storage, and throughput bottlenecks. The ideal candidate has deep hands-on experience in production AI systems, particularly at the systems layer, and is comfortable working on infrastructure where storage architecture and systems efficiency materially affect AI performance.
Skills this posting asks for
- llm serving
- inference systems
- gpu
- cpu
- kv cache
- memory
- storage
- rag
- retrieval
- distributed systems
- model serving
- caching
- distributed performance
- ai infrastructure
- high-performance systems
- storage platforms
Requirements
- Level: senior
From the employer’s posting
WHAT YOU’LL DO - Build and optimize LLM serving and inference systems for production environments - Improve performance across GPU and CPU pathways - Work on KV cache, memory, storage, and throughput bottlenecks - Design and scale systems that support RAG and retrieval-heavy AI workloads…
Read the full description on DDN’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at DDN
- Principal Product Manager, ExaScalerRemote
- Senior Analytics EngineerRemote
- Senior Software Engineering Manager – KV Cache PlatformRemote
- Staff Security EngineerRemote
- Senior/Staff Fuse DeveloperRemote
- Senior Technical Product ManagerRemote