← Back to results

model-optimization jobs in San Diego

$158,400 – $237,600 · Posted 1 day ago

Build and optimize scalable LLM inference platforms at Qualcomm's Cloud AI team, implementing advanced serving techniques like KV-cache management, speculative algorithms, and model optimization. Contribute to production serving frameworks (vLLM, SGLang, Triton, TGI) and work with customers on deployment solutions. Collaborate with compiler, firmware, and platform teams to drive efficient serving through autoscaling, load balancing, and routing. Requires deep understanding of transformer architectures, strong PyTorch and Python skills, computer architecture knowledge, and hands-on experience profiling and optimizing deep learning workloads.

San DiegoLast seen today
$140,800 – $211,200 · Posted 3 days ago

As a Senior Software Engineer focused on AI Tools, you will reauthor and optimize generative AI models (LLMs like Llama, Phi, Qwen, and multimodal models) for efficient execution on Qualcomm's on-device hardware. You'll translate hardware constraints into model-level transformations that preserve accuracy while enabling edge deployment, integrate inference acceleration techniques, and collaborate with compiler and quantization teams to move research prototypes into production. The role requires deep implementation-level knowledge of generative AI architectures, strong Python proficiency in large typed codebases, and hands-on experience optimizing inference for resource-constrained environments.

San DiegoLast seen 1 day ago
Posted 5 days ago

Conduct cutting-edge research in efficient generative AI, including LLMs, multi-modal foundation models, LLM reasoning, reinforcement learning, and agentic AI. Lead high-impact research initiatives and develop production-ready solutions with real-world commercial impact, with a focus on on-device deployment challenges. Implement and evaluate solutions in both simulation and on-device environments, and publish findings in top-tier AI/ML conferences. Requires a Master's degree and 2+ years of experience (or PhD), with proven research excellence, deep expertise in generative AI and reinforcement learning, and hands-on experience with model development pipelines including training, fine-tuning, evaluation, and optimization.

San DiegoLast seen 4 days ago
$179,200 – $268,800 · Posted 7 days ago

As a Staff Product Data Manager for AI Platforms & Systems Solutions at Qualcomm, you will own the product strategy, definition, and lifecycle for AI platform products and specifications spanning model enablement, agentic systems, orchestration runtimes, and open protocols. You will develop comprehensive plans of record including schedules, budgets, and resources, directing products from conception through end of life while functioning as a central resource across systems, engineering, standards, quality, marketing, and partner teams. The role requires deep fluency in the Windows edge-AI stack, with technical depth across AI/ML models, agentic systems, orchestration engines, standards development, and hands-on experience delivering Windows-based AI products leveraging local accelerators (CPU, GPU, NPU). You will translate emerging model, agentic, and orchestration trends into differentiated product and specification roadmaps.

San DiegoLast seen 4 days ago
$140,800 – $211,200 · Posted 8 days ago

Design, develop, and optimize machine learning systems for production AI platforms, including model development, inference optimization, and scalable ML infrastructure. Build training-to-deployment pipelines, optimize model serving for latency and cost, and integrate LLMs and generative AI models into microservices and APIs. Develop data pipelines for ingestion, preprocessing, and feature engineering while collaborating cross-functionally with product, platform, and hardware teams to deliver end-to-end ML solutions.

San DiegoLast seen 7 days ago
$140,800 – $211,200 · Posted 10 days ago

Qualcomm seeks a Machine Learning R&D leader to design and deliver Agentic AI systems for hardware design automation. The role combines hands-on technical execution with team leadership, requiring deep expertise in generative AI, large language models, and ML/AI product development. Responsibilities include architecting end-to-end ML solutions, coaching an interdisciplinary team of researchers and engineers, and bridging research prototypes into production design tools. The ideal candidate brings PhD-level credentials, 2+ years of industry ML/GenAI experience, and proven expertise in agentic AI frameworks, foundation models, and complex ML software pipelines.

San DiegoLast seen 8 days ago
$122,800 – $184,200 · Posted 11 days ago

As an AI Software Engineer at Qualcomm, you will develop and optimize machine learning software for the Qualcomm AI Stack (QAIRT and Genie SDKs), enabling efficient execution of large language models and deep neural networks on Snapdragon platforms. You will validate and optimize performance of generative AI models, debug complex issues, and collaborate across teams to deliver robust AI solutions. The role requires strong software design, programming, and debugging skills, with expertise in machine learning frameworks, embedded systems optimization, and languages like Python or C++. You will work with minimal supervision, participate in design reviews, and contribute to a culture of technical excellence.

San DiegoLast seen 9 days ago