ModularInternal deployment & tooling
Inference Optimization Engineer
Los Altos, CA · US remoteRemote
This role is for an Inference Optimization Engineer on the Performance Labs team at Modular (a Qualcomm company), focused on optimizing LLM inference performance on Modular Cloud across GPU and ASIC architectures.
$198,000–$286,000
View job →