Staff Machine Learning Engineer – AI/ML Compiler
San DiegoLast seen 7 days ago
Summary
Design and develop the end-to-end compilation pipeline for Qualcomm AI Hub Workbench, enabling PyTorch and ONNX models to compile into deployable artifacts targeting LiteRT, ONNXRuntime, and QAIRT across Snapdragon SoCs. Own graph optimization, backend dispatch across CPU/GPU/NPU, model catalog validation, and automated CI/CD pipelines for on-device ML deployment. Requires 4+ years of software/hardware/systems engineering experience and proficiency in Python and C++, with preferred expertise in ML compiler infrastructure, MLIR/ONNX/TVM, torch.export, and on-device frameworks.