← Back to results

triton jobs in San Diego

$200,000 – $250,000 · Posted 20 days ago

AppFolio is hiring a Staff Machine Learning Engineer to design, build, and operate their ML platform on AWS, supporting training, fine-tuning, inference, RAG, and cost optimization across the organization's AI initiatives. You'll partner with applied AI and research teams to productionize prototypes, maintain multi-provider LLM reliability (OpenAI, Google, Anthropic), and operate AI safety guardrails and authorization layers.… The role requires production-scale ML infrastructure experience on AWS (ECS, SageMaker, GPU fleets), deep knowledge of model serving and inference optimization, hands-on language model training, and demonstrated cost discipline across AI workloads.

San DiegoLast seen 19 days ago
$158,400 – $237,600 · Posted 1 month ago

Lead end-to-end model optimization and transformation for large language models, vision language models, diffusion, and multimodal models on Qualcomm inference accelerators. Architect PyTorch-based optimization strategies, drive graph capture and deployment using PyTorch/ONNX/torch.compile, and design fusion kernels using Triton or similar DSLs.… Partner with compiler, performance, and accuracy teams to co-design lowering strategies, optimize transformer-specific patterns (KVcache, decoding, long context), and scale distributed inference across multi-core and multi-device systems. Requires expert-level PyTorch proficiency, deep transformer architecture knowledge, hands-on experience with torch.compile/TorchDynamo, and strong foundation in ML accelerators and distributed systems.

San DiegoLast seen 28 days ago
$200,000 – $250,000 · Posted 1 month ago

Lead machine learning strategy and development for AppFolio's Leasing products, owning the ML roadmap and autonomous leasing agent architecture. Build evaluation frameworks, model quality infrastructure, and establish ML standards across the Leasing Engineering team while ensuring production-grade reliability, SLOs, and observability.… Translate research into shipped features by evaluating fine-tuning approaches, RAG patterns, and agentic systems; operate with production discipline on a SaaS platform serving real customer workflows.

San DiegoLast seen 14 days ago
$248,000 – $372,000 · Posted 1 month ago

Lead Qualcomm's next-generation AI accelerator platform as senior engineering director, defining software vision, architecture, and roadmap while managing multiple engineering organizations spanning system software, device drivers, AI runtimes, compilers, ML frameworks, and platform SDKs. Partner with hardware teams on co-design, drive software readiness across chip development milestones, and establish performance/power/reliability goals.… Requires 9+ years software engineering experience (or PhD + 8 years), 4+ years with C/C++/Java/Python, and demonstrated expertise in AI/ML stacks, compiler technologies, runtime systems, or high-performance computing.

San DiegoLast seen 1 month ago
$248,000 – $372,000 · Posted 1 month ago

Sr. Director leading Qualcomm's next-generation AI accelerator platform software strategy, architecture, and delivery.… Manages multiple engineering organizations spanning system software, firmware, device drivers, AI runtimes, compilers, ML frameworks, SDKs, and validation. Defines performance and reliability goals, drives cross-functional hardware-software co-design, and partners with customers and hyperscalers. Requires 15+ years software engineering experience with 8+ years leading large, globally distributed teams; deep expertise in AI/ML stacks, compilers, runtime systems, or heterogeneous computing architectures.

San DiegoLast seen 1 month ago
$178,400 – $267,600 · Posted 1 month ago

Qualcomm is seeking an AI Performance Engineer to optimize and deploy machine learning models for inference acceleration on cloud AI hardware. The role involves converting and optimizing language models, vision models, and diffusion models using PyTorch and ONNX, analyzing performance bottlenecks, and designing efficient kernels in Triton.… You will collaborate with compiler, firmware, and platform teams to map next-generation AI workloads onto hardware, working across the full product lifecycle from research to commercial deployment. The position requires deep expertise in transformer architectures, inference optimization techniques, computer architecture, ML accelerators, and distributed systems.

San DiegoLast seen 5 days ago