Unknown Company

Staff Machine Learning Engineer – AI/ML Compiler

san diego, ca • Posted 6 days ago
Onsite Full Time General

Qualcomm AI Hub Compiler Team PositionQualcomm AI Hub is the platform for on-device AI — enabling developers to easily integrate, optimize, and deploy ML models on Qualcomm devices. Qualcomm AI Hub Workbench lets developers compile trained PyTorch or ONNX models into deployable artifacts targeting a variety of runtimes — LiteRT, ONNXRuntime, or Qualcomm AI Engine Direct SDK (QAIRT) — and profile and validate them on real Qualcomm devices hosted in the cloud.Join the Qualcomm AI Hub Compiler team and own the infrastructure that powers these model compilations. You will work across the full compilation pipeline — from model ingestion and graph optimization to backend dispatch across CPU, GPU, and NPU — ensuring models compile correctly, execute efficiently, and scale across a growing catalog of on-device use cases spanning vision, audio, speech, and multi-modal models.Compiler Pipeline & InfrastructureDesign, develop, and maintain the end-to-end compilation pipeline powering Qualcomm AI Hub Workbench, from PyTorch and ONNX model ingestion through graph optimization to deployable artifacts targeting LiteRT, ONNXRuntime, or QAIRT on Snapdragon SoCsBuild and maintain ONNX-based compilation paths using ONNX IR: graph transformation passes, op validation, and opset compatibility handlingBuild and maintain PyTorch compilation paths consuming torch.export output, including dynamic shapes, custom ops, and ATen IR decompositionContribute to ONNXRuntime QNN execution provider: graph optimizations, graph partitioning, and op validation and loweringsCollaborate with QAIRT and QNN teams to ensure correct and efficient model execution across CPU, GPU, and NPU backendsBuild tooling to analyze, profile, and debug compilation failures, accuracy regressions, and performance degradations; develop clear, actionable developer-facing diagnosticsModel Catalog, Automation & CollaborationOwn compilation and validation of models published on Qualcomm AI Hub, ensuring correct conversion and verified performance across supported runtime targetsBuild and maintain automated compilation pipelines and CI/CD evaluation harnesses to scale model onboarding as the Qualcomm AI Hub model catalog growsPartner with internal Business Units to onboard models through Qualcomm AI Hub compilation workflows, translating deployment constraints (target SoC, latency budgets, memory limits) into concrete compilation strategiesAuthor technical documentation, tutorials, and example notebooks for the Qualcomm AI Hub developer communityMinimum Qualifications:• Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 4+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.

OR Master's degree in Computer Science, Engineering, Information Systems, or related field and 3+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. OR PhD in Computer Science, Engineering, Information Systems, or related field and 2+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.Preferred Qualifications:3+ years of industry experience in ML infrastructure, compiler engineering, or AI framework developmentProficient in Python and C++Solid understanding of ML compiler concepts (graph IRs, operator fusion, shape inference, lowering passes, backend partitioning) and hands-on experience with one or more compiler stacks such as MLIR, ONNX, or TVMExperience with PyTorch model export (torch.export, torch.compile, FX, ATen IR) and on-device deployment frameworks such as LiteRT, ExecuTorch, or ONNXRuntimeFamiliarity with SoC-level constraints (memory bandwidth, compute precision, NPU/DSP execution) and hardware-specific runtimes such as QAIRT/QNN is a plusExperience building automated CI/CD pipelines for model compilation and validation at scaleStrong written and verbal communication skills; proficiency with git and software engineering best practicesLevel of ResponsibilityWorks independently on open-ended compiler and infrastructure challengesProvides technical guidance and mentorship to team membersDecision-making has broad impact — affecting compilation correctness, runtime performance, and the developer experience across Qualcomm AI HubCommunicates complex compiler and runtime concepts to varied audiences: SoC engineers, BU partners, and external ML developersHas meaningful influence on the Qualcomm AI Hub compiler roadmap, model catalog strategy, and cross-team runtime integration prioritiesPay range and Other Compensation & Benefits :$160,500.00 - $240,700.00The above pay scale reflects the broad, minimum to maximum, pay scale for this job code for the location for which it has been posted. Even more importantly, please note that salary is only one component of total compensation at Qualcomm.

We also offer a competitive annual discretionary bonus program and opportunity for annual RSU grants (employees on sales-incentive plans are not eligible for our annual bonus). In addition, our highly competitive benefits package is designed to support your success at work, at home, and at play. Your recruiter will be happy to discuss all that Qualcomm has to offer – and you can review more details about our US benefits at this link.

Back to Job Search