← Back to results

transformers jobs in San Diego

$158,400 – $237,600 · Posted 1 day ago

Build and optimize scalable LLM inference platforms at Qualcomm's Cloud AI team, implementing advanced serving techniques like KV-cache management, speculative algorithms, and model optimization. Contribute to production serving frameworks (vLLM, SGLang, Triton, TGI) and work with customers on deployment solutions. Collaborate with compiler, firmware, and platform teams to drive efficient serving through autoscaling, load balancing, and routing. Requires deep understanding of transformer architectures, strong PyTorch and Python skills, computer architecture knowledge, and hands-on experience profiling and optimizing deep learning workloads.

San DiegoLast seen today
$140,800 – $211,200 · Posted 3 days ago

As a Senior Software Engineer focused on AI Tools, you will reauthor and optimize generative AI models (LLMs like Llama, Phi, Qwen, and multimodal models) for efficient execution on Qualcomm's on-device hardware. You'll translate hardware constraints into model-level transformations that preserve accuracy while enabling edge deployment, integrate inference acceleration techniques, and collaborate with compiler and quantization teams to move research prototypes into production. The role requires deep implementation-level knowledge of generative AI architectures, strong Python proficiency in large typed codebases, and hands-on experience optimizing inference for resource-constrained environments.

San DiegoLast seen 1 day ago
$162,600 – $244,000 · Posted 7 days ago

As a Datacenter AI Systems and Solutions Engineer at Qualcomm, you will research, develop, and optimize end-to-end AI/ML solutions that integrate Qualcomm's AI inference accelerators with system software and ecosystem components. You will lead the design and deployment of production-ready Generative AI and LLM applications, perform model benchmarking and performance analysis, and serve as a technical lead for customer engagements on AI model optimization and inference tuning. The role requires deep expertise in AI systems architecture, MLOps practices, and large-scale distributed systems, with hands-on proficiency in Python, ML frameworks, containerization, and orchestration platforms. You will drive system-level requirements, hardware/software co-design, and influence product direction through performance analysis and optimization strategies.

San DiegoLast seen 4 days ago
$111,300 – $166,900 · Posted 8 days ago

As a Senior Systems Engineer for Data Center AI at Qualcomm, you will research, develop, optimize, and validate AI/ML solutions that integrate Qualcomm's hardware accelerators with software and ecosystem to deliver high-performance inference in datacenter environments. You will design and deploy Gen AI and LLM applications, implement fine-tuning and distillation techniques, perform model benchmarking, and drive system-level architecture decisions. You will collaborate across functional teams to meet system-level requirements while staying current with advances in AI/ML models and hardware. Required skills include strong Python proficiency, expertise in ML frameworks (PyTorch, TensorFlow), deep understanding of ML deployment and system performance profiling, and experience with parallel computing.

San DiegoLast seen 6 days ago
$121,625 – $217,711 · Posted 10 days ago

The Applied AI Engineer develops, deploys, and maintains generative AI solutions for insurance products and internal workflows. The role involves implementing end-to-end AI pipelines using AWS services (SageMaker, Lambda, ECS/EKS, S3), building feature engineering workflows in Snowflake, and collaborating with cloud and MLOps teams. Candidates need 3+ years of AI/ML engineering experience with 1–2 years focused on generative AI or LLMs, proficiency in Python and frameworks like PyTorch or Hugging Face Transformers, and hands-on cloud deployment experience. The position requires monitoring model performance, ensuring security and compliance, and supporting AI architecture reviews.

San DiegoLast seen 8 days ago