Machine Learning Performance Engineer - Offboard Training & Inference

Applied · Sunnyvale

Posted
11 days ago
Last confirmed live
1 day ago

What this role involves

Applied Intuition is hiring a Machine Learning Performance Engineer to optimize distributed training and offline inference workloads in the datacenter. The role focuses on improving throughput, cluster goodput, and cost efficiency for large-scale ML workloads. The position requires working onsite at their office 5 days a week.

Skills this posting asks for

  • machine learning
  • distributed training
  • inference
  • profiling
  • performance optimization
  • data loading
  • preprocessing
  • augmentation
  • kernel execution
  • gradient communication
  • checkpointing
  • batching
  • scheduling
  • quantization
  • low-precision execution
  • graph optimization
  • accelerator saturation
  • roofline models
  • performance models
  • multi-node scaling
  • sharding
  • parallelism
  • collective communication
  • interconnect utilization

Requirements

  • Remote policy: onsite

From the employer’s posting

Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, tru…

Read the full description on Applied’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at Applied

All 54 roles at Applied