← Back to results

distributed-training jobs in San Diego

$126,700 – $217,900 · Posted 18 days ago

Qualcomm seeks a Senior AI Performance Architect to design hardware accelerator architectures optimized for AI training at scale. The role involves analyzing GPU/accelerator architectures, defining computational blocks supporting multiple datatypes and precisions, architecting memory and scale-out systems, and codesigning hardware with software/LLM requirements.… Responsibilities include competitive analysis, performance modeling, and pre-silicon prediction for ML training workloads. The ideal candidate spans software architecture, algorithm development, kernel optimization, and hardware accelerator design.

San DiegoLast seen 16 days ago
Posted 19 days ago

Design, build, and operate production ML systems across the full lifecycle—from data preparation and model training through deployment, monitoring, and continuous improvement. Partner with data scientists, software engineers, and platform teams to turn research prototypes into reliable, scalable services.… Work across recommendation, ranking, forecasting, classification, NLP, and generative AI depending on product priorities. Balance model quality, inference latency, scalability, and operational resilience as equally important outcomes.

San DiegoLast seen 17 days ago
$233,000 – $350,000 · Posted 1 month ago

Senior Staff Engineer responsible for designing and building an MLOps platform that supports distributed AI training, reinforcement learning, and foundation model development at scale. You will architect Kubernetes-native infrastructure for GPU workloads, design self-service AI development workflows, manage the data and model lifecycle, and lead platform distribution across cloud, on-premises, and air-gapped environments.… The role requires deep expertise in modern AI frameworks (PyTorch, Hugging Face Transformers), distributed systems, GPU scheduling, and cloud-native infrastructure, with a focus on enabling researchers and engineers to move from experimentation to production rapidly.

San DiegoLast seen 1 month ago
Posted 1 month ago

As Principal Software Engineer for AI & Data Platform, you will architect the data foundation for scientific and engineering R&D platforms, designing scalable data processing patterns, ML training pipelines, and intelligent workflow interfaces. You will own end-to-end responsibilities including data modeling for multi-use analytics and ML, building production training and fine-tuning pipelines, model evaluation and benchmarking, and setting engineering standards for the team.… The role requires 10+ years shipping production software, expert-level Python, deep experience with large-scale data systems (object storage, analytical processing, training formats), hands-on ML pipeline development, and the ability to set technical direction in early-stage environments while implementing it yourself.

San DiegoLast seen 18 days ago
$200,900 – $257,500 · Posted 1 month ago

Design and build large-scale multimodal foundation models that integrate high-dimensional biomedical imaging with molecular and language data. You will implement cross-domain data fusion strategies, develop high-performance ML pipelines for petabyte-scale datasets, and collaborate with scientists and engineers to translate biological complexity into scalable, production-ready code.… The role requires deep expertise in Computer Vision architectures, distributed training frameworks, and expert Python proficiency. You will contribute through high-impact publications or open-source work and demonstrated ability to handle the noise and sparsity inherent in biological data.

San DiegoLast seen 1 month ago