← Back to results

sagemaker jobs in San Diego

Posted 4 days ago

This role is a technical cloud architect for AWS Professional Services focused on Life Sciences and Manufacturing/Supply Chain customers. You will design and implement complex cloud solutions involving AI/ML, IoT, digital twins, supply chain optimization, and manufacturing data platforms — including CMC orchestration, MES/MOM modernization, and agentic supply chain deployments.… You'll drive solution architecture from prospecting through sales closure, translating regulatory and compliance requirements (cGMP, GxP, 21 CFR Part 11, FDA/EMA) into technical designs using AWS services like IoT SiteWise, IoT TwinMaker, Bedrock, SageMaker, and Supply Chain tools. The role requires 8+ years of cloud architecture, 7+ years in Healthcare/Life Sciences or Manufacturing industries, and deep expertise in Data & AI technologies, with a bachelor's degree in computer science, engineering, or a quantitative field.

San DiegoLast seen 3 days ago
$150,000 – $230,000 · Posted 14 days ago

Staff Engineer responsible for designing and building scalable test automation frameworks and CI/CD infrastructure for machine-learning systems at scale. The role requires expertise validating ML quality across the full lifecycle (data, training, evaluation, packaging, deployment, monitoring), GPU-accelerated workloads in Kubernetes, and integrated hardware-software systems.… You will develop automated evaluation suites, performance baselines, release thresholds, and observability solutions while mentoring teams on ML testing best practices and dependency/reproducibility standards.

San DiegoLast seen 12 days ago
$150,000 – $230,000 · Posted 14 days ago

Staff Engineer responsible for designing and implementing scalable test automation frameworks and infrastructure for machine learning systems across multiple engineering teams. The role requires expertise in validating ML pipelines across data, training, evaluation, deployment, and monitoring; testing GPU-accelerated workloads in Kubernetes; qualifying integrated hardware and software systems; and developing comprehensive integration and regression strategies.… Deep experience with Python automation, distributed systems testing, performance profiling, and observability is essential, along with strong system-design skills for complex, multi-tenant environments.

San DiegoLast seen 13 days ago
$247,500 – $335,000 · Posted 14 days ago

Principal Engineer leading AI-native application development across Intuit's products (TurboTax, Credit Karma). Drives end-to-end technology initiatives, designs scalable distributed systems, and applies AI/LLM technologies to customer problems.… Requires 8+ years developing software for large enterprises, 5+ years designing complex distributed systems, full-stack development experience with AI tools, and proficiency across front-end (React, Angular, SwiftUI, Kotlin), back-end (Java, TypeScript, Spring, Express), and cloud platforms (AWS, GCP).

San DiegoLast seen 12 days ago
$200,000 – $250,000 · Posted 24 days ago

AppFolio is hiring a Staff Machine Learning Engineer to design, build, and operate their ML platform on AWS, supporting training, fine-tuning, inference, RAG, and cost optimization across the organization's AI initiatives. You'll partner with applied AI and research teams to productionize prototypes, maintain multi-provider LLM reliability (OpenAI, Google, Anthropic), and operate AI safety guardrails and authorization layers.… The role requires production-scale ML infrastructure experience on AWS (ECS, SageMaker, GPU fleets), deep knowledge of model serving and inference optimization, hands-on language model training, and demonstrated cost discipline across AI workloads.

San DiegoLast seen 23 days ago
$198,500 – $297,700 · Posted 1 month ago

Lead the design, development, and operation of Qualcomm's enterprise AI platform, supporting agentic AI, model lifecycle management, and multi-cloud ML/inference infrastructure. Own core platform services including identity/RBAC, secrets, service meshes, observability, vector stores, and model gateways across on-prem GPU clusters and managed cloud services.… Manage a ~10-engineer global team (platform, SRE, MLOps/LLMOps), drive incident response and continuous improvement, and partner with product and security teams on AI governance and high-impact use cases.

San DiegoLast seen 20 days ago
$105,780 – $189,348 · Posted 1 month ago

AI Architect role requires designing and delivering enterprise-scale AI/ML solutions on AWS, with 5–7 years of cloud experience and 2–3 years hands-on with LLMs, agentic frameworks, and AI/ML services. The position demands deep expertise in AWS AI services (Bedrock, SageMaker, Kendra), orchestration frameworks (LangChain, LangGraph), RAG pipelines, MLOps tooling, and responsible AI principles.… Candidates must architect scalable, cost-optimized solutions integrating AI with cloud infrastructure, data platforms, and microservices, and communicate complex technical concepts to executive and technical audiences.

San DiegoLast seen 24 days ago
$155,600 – $306,800 · Posted 1 month ago

Senior Forward Deployed Engineer responsible for building and deploying GenAI/LLM-powered solutions on AWS, working directly with enterprise clients to translate business needs into production AI systems. The role requires 5+ years of software/data engineering experience, 1+ years hands-on with AWS AI&Data services (Bedrock, Neptune, OpenSearch), and the ability to lead technical workstreams while mentoring team members.… You will prototype solutions, develop scalable AI patterns with human-in-the-loop controls, write production-quality code with strong testing and CI/CD practices, and create reusable assets including code libraries and reference implementations. 50% travel required.

San DiegoLast seen 1 month ago
Posted 1 month ago

Own and deliver small to medium machine learning system components from design through production deployment, including building data pipelines, training and evaluating models, and implementing MLOps monitoring. Write high-quality Python code to translate technical requirements into maintainable solutions, working with frameworks like scikit-learn, TensorFlow, PyTorch, and HuggingFace.… Design and deploy ML models as microservices, APIs, batch jobs, or streaming components on AWS, with responsibility for model performance metrics, data drift detection, and retraining triggers. Collaborate across Data Engineers, Software Engineers, Data Scientists, and product stakeholders to deliver project objectives.

San DiegoLast seen 1 month ago