← Back to results

cloud-cost-management jobs in San Diego

$109,200 – $150,150 · Posted 7 days ago

The Cloud Optimization Engineer – FinOps will manage hybrid cloud spend across private and public cloud environments, using Flexera to monitor costs and deliver monthly reporting and dashboards to stakeholders. The role involves analyzing cloud architectures and workload patterns to identify cost drivers, building and automating data pipelines and cost models, and implementing cost optimization actions such as rightsizing and reservations.… You will collaborate hands-on with engineering teams, develop forecasting models, enforce tagging and governance standards, and drive FinOps best practices across the organization. The position requires 5+ years of FinOps and cloud financial management experience, proficiency in Azure or GCP, and advanced skills in Excel, SQL, Python, and PowerShell.

San DiegoLast seen 5 days ago
$190,000 – $280,000 · Posted 12 days ago

Senior Staff Lead Site Reliability Engineer to establish and mature SRE practices across Hivemind's cloud infrastructure and platform services. You will define reliability targets (SLIs/SLOs), build observability systems, lead incident response and root-cause analysis, mentor teams on reliability-first practices, and develop automation to reduce manual operational work.… The role requires 7+ years in SRE or infrastructure engineering, hands-on experience operating production services in AWS or equivalent cloud environments, expertise with containerized/distributed systems, infrastructure-as-code, and operational tooling development in Python or Go.

San DiegoLast seen 10 days ago
$161,700 – $242,500 · Posted 17 days ago

The Enterprise AI FinOps Lead establishes financial governance for company-wide AI usage across multi-cloud (Azure, AWS, GCP), SaaS, and on-premise environments. The role owns cost visibility, show-back, optimization, budgeting, forecasting, and stakeholder engagement, with deep responsibility for AI-specific cost drivers such as tokens, models, prompts, embeddings, and RAG systems.… The leader translates telemetry into actionable insights, optimization programs, and executive reporting while managing the full departmental budget. This is a hands-on operator role requiring 8+ years in FinOps, cloud cost management, or enterprise finance, with strong analytical, stakeholder management, and executive communication skills.

San DiegoLast seen 15 days ago
Posted 20 days ago

Shield AI seeks an experienced SRE Lead to establish and mature reliability practices across Hivemind's cloud infrastructure and platform services. This hands-on technical role involves defining SLIs/SLOs, building monitoring and alerting systems, leading incident response, and driving root-cause analysis.… The SRE Lead will mentor teams, develop operational tooling in Python or Go, and partner with product and cloud engineering teams to embed reliability-first practices into system design and infrastructure provisioning.

San DiegoLast seen 18 days ago
$190,000 – $280,000 · Posted 1 month ago

Shield AI seeks an experienced Site Reliability Engineer to establish and mature the SRE function across Hivemind's cloud infrastructure and platform services. This hands-on leadership role involves defining reliability targets (SLIs/SLOs), improving observability and incident response practices, investigating complex failures, and driving adoption of reliability-first engineering practices.… The SRE Lead will mentor teams, develop operational automation tooling, and manage the multi-quarter SRE roadmap while working closely with Cloud Engineering and product teams to enhance system resilience and operational excellence.

San DiegoLast seen 1 month ago
$220,203 – $275,254 · Posted 1 month ago

Lead a team of cloud engineers building the backend platform that powers Brain Corp's global fleet of 30,000+ autonomous mobile robots. Own the platform's reliability, roadmap, and architecture as it scales to support fleet operations and customer-facing applications.… Balance hands-on technical leadership with people management—mentor engineers, drive hiring and onboarding, own production reliability and incident response, and partner with principal/staff engineers on distributed-systems architecture. Navigate tradeoffs between performance, cost, reliability, and scalability while shipping features that robots in the field depend on.

San DiegoLast seen 1 month ago