Staff Machine Learning Engineer – AI/ML Compiler
Teknik, data och digitalt · Data, AI och analys · Maskininlärning · Mjukvaruutveckling · Data engineering
I korthet
Join Qualcomm AI Hub as a Staff Machine Learning Engineer specializing in AI/ML Compilers. You will own the end-to-end compilation pipeline, from model ingestion to backend dispatch across CPU, GPU, and NPU, ensuring efficient and correct execution on Qualcomm devices. This role involves designing, developing, and maintaining compiler infrastructure, collaborating with various teams, and contributing to the Qualcomm AI Hub platform.
Ansvarsområden
- Design, develop, and maintain the end-to-end compilation pipeline powering Qualcomm AI Hub Workbench.
- Build and maintain ONNX-based compilation paths using ONNX IR.
- Build and maintain PyTorch compilation paths consuming torch.export output.
- Contribute to ONNXRuntime QNN execution provider.
- Collaborate with QAIRT and QNN teams for efficient model execution across CPU, GPU, and NPU backends.
- Build tooling to analyze, profile, and debug compilation failures, accuracy regressions, and performance degradations.
- Own compilation and validation of models published on Qualcomm AI Hub.
- Build and maintain automated compilation pipelines and CI/CD evaluation harnesses.
- Partner with internal Business Units to onboard models through Qualcomm AI Hub compilation workflows.
- Author technical documentation, tutorials, and example notebooks for the Qualcomm AI Hub developer community.
Krav
- Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 4+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
- OR Master's degree in Computer Science, Engineering, Information Systems, or related field and 3+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
- OR PhD in Computer Science, Engineering, Information Systems, or related field and 2+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
- Proficient in Python and C++.
- Solid understanding of ML compiler concepts (graph IRs, operator fusion, shape inference, lowering passes, backend partitioning) and hands-on experience with one or more compiler stacks such as MLIR, ONNX, or TVM.
- Experience with PyTorch model export (torch.export, torch.compile, FX, ATen IR) and on-device deployment frameworks such as LiteRT, ExecuTorch, or ONNXRuntime.
- Experience building automated CI/CD pipelines for model compilation and validation at scale.
- Strong written and verbal communication skills; proficiency with git and software engineering best practices.
Önskade kvalifikationer
- 3+ years of industry experience in ML infrastructure, compiler engineering, or AI framework development.
- Familiarity with SoC-level constraints (memory bandwidth, compute precision, NPU/DSP execution) and hardware-specific runtimes such as QAIRT/QNN is a plus.
Förmåner
- Competitive annual discretionary bonus program.
- Opportunity for annual RSU grants.
- Highly competitive benefits package designed to support your success at work, at home, and at play.
- Qualcomm is an equal opportunity employer.
#AI#ML#Compiler#Machine Learning#Engineering