← Back to results

onnx jobs in San Diego

$140,800 – $211,200 · Posted 6 days ago

Design, develop, and optimize machine learning systems for production AI platforms, focusing on model development, inference optimization, and scalable ML infrastructure. Responsibilities include building ML pipelines and frameworks, optimizing model inference for latency and cost, integrating LLMs into APIs and microservices, and designing data pipelines for preprocessing and feature engineering.… The role requires strong software engineering fundamentals combined with deep ML expertise, working across PyTorch/TensorFlow, model-serving systems, distributed computing, and GPU environments.

San DiegoLast seen 4 days ago
$122,800 – $184,200 · Posted 6 days ago

Design, develop, and optimize production machine learning systems including model development, inference optimization, and scalable ML infrastructure. Build end-to-end ML pipelines from training through deployment, optimize model inference for latency and cost, integrate LLM/ML models into APIs and microservices, and engineer data pipelines for preprocessing and validation.… Requires strong software engineering fundamentals with deep ML expertise, proficiency in Python and systems languages, and experience with ML frameworks, model serving, and distributed computing.

San DiegoLast seen 4 days ago
$174,000 – $261,000 · Posted 12 days ago

Design and develop the end-to-end compilation pipeline for Qualcomm AI Hub Workbench, enabling PyTorch and ONNX models to compile into deployable artifacts targeting LiteRT, ONNXRuntime, and QAIRT across Snapdragon SoCs. Own graph optimization, backend dispatch across CPU/GPU/NPU, model catalog validation, and automated CI/CD pipelines for on-device ML deployment.… Requires 4+ years of software/hardware/systems engineering experience and proficiency in Python and C++, with preferred expertise in ML compiler infrastructure, MLIR/ONNX/TVM, torch.export, and on-device frameworks.

San DiegoLast seen 11 days ago
$129,500 – $194,300 · Posted 27 days ago

Qualcomm seeks an ML/computer vision engineer to develop and optimize AI-accelerated vision algorithms for edge devices (mobile, automotive, robotics, AR/VR). You will design algorithms for depth estimation, optical flow, video super-resolution, 3D reconstruction, and SLAM, then optimize deep learning models for resource-constrained SoCs.… The role spans algorithm research, model architecture design, performance profiling across memory/compute/power/bandwidth, and cross-functional collaboration with hardware and systems teams to ship production models for key customer segments.

San DiegoLast seen 24 days ago
$200,800 – $301,200 · Posted 27 days ago

Lead AI software strategy and architecture for Qualcomm's automotive platform, directing deployment of computer vision, perception systems, LLMs, and generative AI workloads. Drive architectural decisions across multiple products and customer programs while mentoring cross-functional teams.… Ensure solutions meet automotive safety, performance, and reliability standards. Requires 8+ years of software/systems engineering experience with deep expertise in AI/ML, C/C++, and automotive software environments.

San DiegoLast seen 26 days ago
$200,800 – $301,200 · Posted 28 days ago

Principal Software Engineer leading AI software strategy and architecture for Qualcomm's automotive platform. Responsibilities include architecting and deploying advanced AI workloads (computer vision, LLMs, multimodal models, generative AI) on automotive systems, driving cross-functional technical decisions across AI, automotive engineering, and platform teams, mentoring engineers organization-wide, and ensuring solutions meet automotive safety, performance, and reliability requirements.… Requires 8+ years of software/systems engineering experience (or equivalent with advanced degree); preferred qualifications include 15+ years in AI/ML, embedded systems, or automotive software, expert C/C++ proficiency, and demonstrated experience with automotive safety standards (ISO 26262, ASPICE), edge AI inference, and Qualcomm AI frameworks.

San DiegoLast seen 27 days ago
$158,400 – $237,600 · Posted 1 month ago

Lead end-to-end model optimization and transformation for large language models, vision language models, diffusion, and multimodal models on Qualcomm inference accelerators. Architect PyTorch-based optimization strategies, drive graph capture and deployment using PyTorch/ONNX/torch.compile, and design fusion kernels using Triton or similar DSLs.… Partner with compiler, performance, and accuracy teams to co-design lowering strategies, optimize transformer-specific patterns (KVcache, decoding, long context), and scale distributed inference across multi-core and multi-device systems. Requires expert-level PyTorch proficiency, deep transformer architecture knowledge, hands-on experience with torch.compile/TorchDynamo, and strong foundation in ML accelerators and distributed systems.

San DiegoLast seen 28 days ago
$160,500 – $240,700 · Posted 1 month ago

Staff Machine Learning Engineer role focused on designing and maintaining Qualcomm AI Hub's end-to-end ML compilation pipeline, from PyTorch and ONNX model ingestion through graph optimization to deployable artifacts on Snapdragon SoCs. Responsibilities span compiler infrastructure (graph transformation, op validation, backend dispatch across CPU/GPU/NPU), model catalog automation and CI/CD, and developer tooling for profiling and debugging.… Requires 4+ years of software/systems engineering experience (or 3+ with a master's, or 2+ with a PhD), with preferred expertise in ML compiler concepts, ONNX/PyTorch export workflows, and on-device deployment frameworks.

San DiegoLast seen 1 month ago
$158,400 – $237,600 · Posted 1 month ago

Staff/Sr. Staff Software Engineer role focused on AI inference optimization on Snapdragon platforms, including model optimization, quantization, graph transformations, and runtime execution for LLMs, LVMs, and LMMs.… You will design and implement graph lowering and optimization techniques within ONNX Runtime, ExecuTorch, and Qualcomm AI Stack SDK, working across ML algorithms, inference systems, and hardware integration. The role requires 6–8+ years of software development experience, 3+ years in AI/ML inference or model optimization, deep expertise in Python and C/C++, PyTorch/ONNX, and transformer architectures. You will mentor junior engineers, drive features end-to-end, and collaborate across ML Research, hardware, product, and QA teams.

San DiegoLast seen 18 days ago
$148,300 – $222,500 · Posted 1 month ago

Design, implement, and optimize video processing and computer vision algorithms for mobile and embedded platforms at scale, from concept through production deployment. Profile and optimize algorithm performance across memory, compute, power, and bandwidth constraints on SoCs with embedded accelerators.… Lead system and algorithm-level architecture, collaborate with cross-functional hardware and software teams, and provide technical mentorship to engineers across experience levels.

San DiegoLast seen 1 day ago
$248,000 – $372,000 · Posted 1 month ago

Lead Qualcomm's next-generation AI accelerator platform as senior engineering director, defining software vision, architecture, and roadmap while managing multiple engineering organizations spanning system software, device drivers, AI runtimes, compilers, ML frameworks, and platform SDKs. Partner with hardware teams on co-design, drive software readiness across chip development milestones, and establish performance/power/reliability goals.… Requires 9+ years software engineering experience (or PhD + 8 years), 4+ years with C/C++/Java/Python, and demonstrated expertise in AI/ML stacks, compiler technologies, runtime systems, or high-performance computing.

San DiegoLast seen 1 month ago
$248,000 – $372,000 · Posted 1 month ago

Sr. Director leading Qualcomm's next-generation AI accelerator platform software strategy, architecture, and delivery.… Manages multiple engineering organizations spanning system software, firmware, device drivers, AI runtimes, compilers, ML frameworks, SDKs, and validation. Defines performance and reliability goals, drives cross-functional hardware-software co-design, and partners with customers and hyperscalers. Requires 15+ years software engineering experience with 8+ years leading large, globally distributed teams; deep expertise in AI/ML stacks, compilers, runtime systems, or heterogeneous computing architectures.

San DiegoLast seen 1 month ago
$140,800 – $211,200 · Posted 1 month ago

As a Senior Software Engineer focused on AI Tools, you will reauthor and optimize generative AI models (LLMs like Llama, Phi, Qwen, and multimodal models) for efficient execution on Qualcomm's on-device hardware. You'll translate hardware constraints into model-level transformations that preserve accuracy while enabling edge deployment, integrate inference acceleration techniques, and collaborate with compiler and quantization teams to move research prototypes into production.… The role requires deep implementation-level knowledge of generative AI architectures, strong Python proficiency in large typed codebases, and hands-on experience optimizing inference for resource-constrained environments.

San DiegoLast seen 24 days ago
Posted 1 month ago

Own CI, build, and release infrastructure for AI software products (ONNX Runtime, ExecuTorch, TFLite/LiteRT delegates) across Linux, Windows, and Android. Design and evolve multi-platform CI pipelines, implement HIL and on-device automated testing for Snapdragon SoC validation, and drive end-to-end release pipelines with versioning, artifact packaging, and distribution.… Define automated quality gates including regression benchmarks and performance thresholds. Requires 8+ years of software/DevOps/release engineering experience (or 6+ with a Master's), expert Python and Bash scripting, deep CI/CD platform expertise (Jenkins, GitLab CI, GitHub Actions, TeamCity), and hands-on experience integrating AI coding agents into workflows.

San DiegoLast seen 1 month ago
$178,400 – $267,600 · Posted 1 month ago

Qualcomm is seeking an AI Performance Engineer to optimize and deploy machine learning models for inference acceleration on cloud AI hardware. The role involves converting and optimizing language models, vision models, and diffusion models using PyTorch and ONNX, analyzing performance bottlenecks, and designing efficient kernels in Triton.… You will collaborate with compiler, firmware, and platform teams to map next-generation AI workloads onto hardware, working across the full product lifecycle from research to commercial deployment. The position requires deep expertise in transformer architectures, inference optimization techniques, computer architecture, ML accelerators, and distributed systems.

San DiegoLast seen 5 days ago
$140,800 – $211,200 · Posted 1 month ago

Design, develop, and optimize machine learning systems and models for production AI platforms, with focus on inference optimization, scalable ML infrastructure, and deployment. Responsibilities include building ML pipelines, optimizing model inference across hardware environments, integrating LLMs and models into APIs and microservices, designing data pipelines for ingestion and feature engineering, and collaborating cross-functionally on end-to-end ML solutions.… Requires strong software engineering fundamentals combined with deep ML expertise, proficiency in Python and at least one systems language (C++, Rust, or Go), and solid understanding of ML frameworks, transformer architectures, and model deployment systems.

San DiegoLast seen 29 days ago
$140,800 – $211,200 · Posted 1 month ago

Design, develop, and optimize machine learning systems for production AI platforms, including model development, inference optimization, and scalable ML infrastructure. Build training-to-deployment pipelines, optimize model serving for latency and cost, and integrate LLMs and generative AI models into microservices and APIs.… Develop data pipelines for ingestion, preprocessing, and feature engineering while collaborating cross-functionally with product, platform, and hardware teams to deliver end-to-end ML solutions.

San DiegoLast seen 1 month ago