← Back to results

ml-accelerators jobs in San Diego

$158,400 – $237,600 · Posted 1 day ago

Build and optimize scalable LLM inference platforms at Qualcomm's Cloud AI team, implementing advanced serving techniques like KV-cache management, speculative algorithms, and model optimization. Contribute to production serving frameworks (vLLM, SGLang, Triton, TGI) and work with customers on deployment solutions. Collaborate with compiler, firmware, and platform teams to drive efficient serving through autoscaling, load balancing, and routing. Requires deep understanding of transformer architectures, strong PyTorch and Python skills, computer architecture knowledge, and hands-on experience profiling and optimizing deep learning workloads.

San DiegoLast seen today
$140,800 – $211,200 · Posted 1 day ago

Senior Machine Learning Engineer responsible for architecting, designing, developing, and deploying on-device AI prototype software for edge computing applications. The role requires expert-level proficiency in C++ and Python, deep knowledge of ML frameworks like PyTorch, and hands-on experience with low-level system debugging, performance tuning, and native inference driver development. You will work in a multi-disciplinary research team advancing generative AI technology for the edge, including model fine-tuning, hardware acceleration, model quantization, and edge inference. A strong theoretical background in deep learning combined with embedded software development experience and modern software engineering best practices is essential.

San DiegoLast seen today
$178,400 – $267,600 · Posted 6 days ago

Qualcomm is seeking an AI Performance Engineer to optimize and deploy machine learning models for inference acceleration on cloud AI hardware. The role involves converting and optimizing language models, vision models, and diffusion models using PyTorch and ONNX, analyzing performance bottlenecks, and designing efficient kernels in Triton. You will collaborate with compiler, firmware, and platform teams to map next-generation AI workloads onto hardware, working across the full product lifecycle from research to commercial deployment. The position requires deep expertise in transformer architectures, inference optimization techniques, computer architecture, ML accelerators, and distributed systems.

San DiegoLast seen 4 days ago
$122,800 – $184,200 · Posted 6 days ago

Join Qualcomm AI Research as a Machine Learning Engineer to architect, design, develop, and deploy on-device AI prototype software for edge devices. You will work with a multi-disciplinary team using cutting-edge AI frameworks, focusing on model fine-tuning, quantization, and edge inference. Strong Python, deep learning frameworks (PyTorch), object-oriented design, and software engineering best practices are required. Preferred qualifications include generative AI knowledge, neural network optimization, NPU/ML accelerator experience, Android programming, C/C++, and embedded software skills.

San DiegoLast seen 4 days ago