← Back to results

error-budgets jobs in San Diego

Posted 11 days ago

The Site Reliability Engineer owns the systems and practices that keep production services available, performant, and recoverable at scale, focusing on Kubernetes-based workloads, cloud infrastructure, observability, and incident response. The role requires designing and operating highly available infrastructure on AWS or GCP using Kubernetes, Terraform, and infrastructure-as-code; building observability across services using Prometheus, Grafana, OpenTelemetry, and centralized logging; and automating deployment, scaling, backup, and recovery workflows through CI/CD pipelines.… The engineer will define SLIs and SLOs, lead incident response and root-cause analysis, harden systems through access controls and disaster recovery testing, and partner with application teams to improve service design and operational readiness. The position requires 3–8 years of DevOps or platform engineering experience, hands-on Kubernetes and containerized services operation, strong Linux and networking fundamentals, proficiency with Terraform or equivalent infrastructure-as-code, observability implementation experience, and strong scripting or programming skills in Python, Go, or Bash.

San DiegoLast seen 9 days ago
Posted 26 days ago

Design and build platform engineering infrastructure, tools, and automation for cloud-native environments on AWS, using Infrastructure as Code (Terraform/CDK), CI/CD, and observability practices. Develop reusable modules, APIs, CLIs, and service templates that improve developer experience and enforce security, compliance, and cost controls.… Establish monitoring, logging, alerting, SLOs, and incident response practices; participate in 24x7 on-call rotation with a focus on production reliability and automation. Requires 7+ years in DevOps, SRE, platform engineering, or cloud operations with deep hands-on experience in multi-account AWS setups, GitOps, and observability platforms like Datadog or Prometheus.

San DiegoLast seen 24 days ago
$150,000 – $180,000 · Posted 1 month ago

The Systems Engineer applies interdisciplinary expertise to define, develop, integrate, and test complex In Vitro Diagnostic systems for medical device applications. Responsibilities include documenting user and system requirements, decomposing system architecture into subsystems with defined interfaces, conducting hazard and risk analysis, managing configuration and change control, driving system integration and root-cause analysis, and supporting clinical validation.… The role requires 10+ years of industry experience with demonstrated mastery of the full product development lifecycle in medical device or regulated environments, including risk management, design reviews, and cross-functional collaboration with development, testing, and clinical teams.

San DiegoLast seen 1 month ago