← Back to results

pulumi jobs in San Diego

$138,400 – $173,000 · Posted 1 day ago

Sr. Site Reliability Engineer responsible for building and improving observability infrastructure, monitoring systems, and reliability patterns across AppFolio's Real Estate Platform. You'll work with engineering teams to implement SLIs/SLOs, diagnose performance issues across the full stack, and manage infrastructure-as-code deployments on Kubernetes and AWS. Strong coding skills (Go, Ruby, or Python) and 5+ years of industry experience required; you'll be on-call and expected to help teams become self-sufficient in reliability practices.

San DiegoLast seen today
$220,203 – $275,254 · Posted 10 days ago

Lead a team of cloud engineers building the nervous system for Brain Corp's fleet of 30,000+ autonomous mobile robots. Own the platform's reliability, roadmap, and architecture while managing team hiring, onboarding, and career development. Drive high-performance delivery across robot and customer-facing interfaces, balance competing priorities (performance, cost, reliability, scalability), and partner with principal and staff engineers on long-term platform strategy. Stay hands-on enough to make sound architectural decisions and earn senior engineer trust during a period of significant growth.

San DiegoLast seen 7 days ago
$220,203 – $275,254 · Posted 10 days ago

Lead a team of cloud engineers building the backend platform that powers Brain Corp's global fleet of 30,000+ autonomous mobile robots. Own the platform's reliability, roadmap, and architecture as it scales to support fleet operations and customer-facing applications. Balance hands-on technical leadership with people management—mentor engineers, drive hiring and onboarding, own production reliability and incident response, and partner with principal/staff engineers on distributed-systems architecture. Navigate tradeoffs between performance, cost, reliability, and scalability while shipping features that robots in the field depend on.

San DiegoLast seen 7 days ago
$187,363 – $265,900 · Posted 11 days ago

Design and architect next-generation ML inference infrastructure for globally distributed, multi-tenant model serving with high availability, scalability, and cost efficiency. Lead the development of low-latency, high-throughput inference systems supporting computer vision and multimodal models (CNNs, segmentation, object detection) using Scala, Java, and Go. Build large-scale distributed systems with reactive frameworks, integrate enterprise feature stores, and extend CI/CD pipelines (GitHub Prow, Pulumi) with automation and policy enforcement. Design advanced observability frameworks, optimize ML algorithms for performance, mentor engineers, and ensure MLOps and compliance standards (GDPR, SOC2) across the platform.

San DiegoLast seen 9 days ago