← Back to results

argo-cd jobs in San Diego

$150,000 – $200,000 · Posted 5 days ago

Cloud Engineer III will deploy, configure, administer, and troubleshoot cloud-based compute, storage, networking, and application infrastructure on AWS and Kubernetes. The role requires hands-on expertise in Linux systems administration, containerized applications (Docker/Kubernetes), cloud networking, IAM, security hardening, and observability platforms.… Responsibilities include supporting production Kubernetes clusters (EKS), implementing Infrastructure as Code (Terraform/CloudFormation), integrating CI/CD pipelines, automating administration with Python/Bash, and troubleshooting complex issues across cloud, container, and application layers. The position demands strong security practices, compliance knowledge for government environments, and collaboration with cross-functional technical and operational teams.

San DiegoLast seen 3 days ago
$79,040 – $120,640 · Posted 5 days ago

Design, build, and operate secure, scalable DevOps platforms at enterprise scale, with deep expertise in Kubernetes, Jenkins, Docker, and CI/CD infrastructure. Manage containerized workloads, implement monitoring and observability, provision and operate AWS infrastructure, and ensure platform reliability through incident response and capacity planning.… The role requires 3+ years of IT experience (or 5+ without a degree) with strong Linux administration, hands-on Kubernetes and Docker expertise, Jenkins infrastructure administration at scale, and CI/CD platform knowledge. This is a hands-on engineering role focused on platform reliability, security hardening, and developer enablement.

San DiegoLast seen 3 days ago
$120,000 – $135,000 · Posted 9 days ago

Senior technical leader responsible for integrating AI/ML systems into secure, scalable DoD-compliant environments. Will architect AI system deployments, establish DevSecOps practices, ensure compliance with RMF/NIST 800-53/DISA STIG standards, and provide technical guidance across engineering teams.… Requires deep expertise in AI/ML production systems, containerization, CI/CD, and cybersecurity controls, with hands-on experience in Python, Kubernetes, and modern infrastructure tooling.

San DiegoLast seen 7 days ago
Posted 12 days ago

The Site Reliability Engineer owns the systems and practices that keep production services available, performant, and recoverable at scale, focusing on Kubernetes-based workloads, cloud infrastructure, observability, and incident response. The role requires designing and operating highly available infrastructure on AWS or GCP using Kubernetes, Terraform, and infrastructure-as-code; building observability across services using Prometheus, Grafana, OpenTelemetry, and centralized logging; and automating deployment, scaling, backup, and recovery workflows through CI/CD pipelines.… The engineer will define SLIs and SLOs, lead incident response and root-cause analysis, harden systems through access controls and disaster recovery testing, and partner with application teams to improve service design and operational readiness. The position requires 3–8 years of DevOps or platform engineering experience, hands-on Kubernetes and containerized services operation, strong Linux and networking fundamentals, proficiency with Terraform or equivalent infrastructure-as-code, observability implementation experience, and strong scripting or programming skills in Python, Go, or Bash.

San DiegoLast seen 10 days ago
Posted 12 days ago

The DevOps Engineer owns the design, deployment, and operation of production infrastructure at scale across AWS and Kubernetes, with responsibility for CI/CD pipelines, observability, and incident response. The role requires hands-on expertise in AWS services (EC2, EKS, VPC, IAM, S3, RDS), Kubernetes cluster operations, Terraform-based infrastructure-as-code, and monitoring tools like Prometheus and Grafana.… You will automate operational workflows with Python, Go, or Bash, lead incident investigations, and partner with software engineers and SREs to improve release velocity and system reliability. 3–8 years of DevOps, SRE, or platform engineering experience is expected.

San DiegoLast seen 10 days ago
$143,000 – $215,000 · Posted 13 days ago

Senior Platform Engineer to design, build, and operate a cloud-native compute platform on AWS and Kubernetes. You will own platform security, scalability, and evolution through strategic initiatives and migration efforts, lead incident response and troubleshooting, and provide technical leadership and mentorship.… Requires 8+ years in platform/SRE/infrastructure engineering with deep AWS expertise, production Kubernetes experience, hands-on CI/CD (GitHub, GitHub Actions, Argo CD), Infrastructure as Code (Terraform/CDK), Kubernetes networking knowledge, and production observability/monitoring skills.

San DiegoLast seen 11 days ago
$175,000 – $195,000 · Posted 13 days ago

Senior or Staff Cloud Infrastructure Engineer responsible for designing and operating Crucible, a manufacturing-network software platform running across commercial cloud, government cloud, and on-premises environments. You will own infrastructure-as-code (Terraform, Helm, GitOps), Kubernetes production operations, cloud networking, identity and security controls, CI/CD pipelines, and observability tooling.… You'll work directly with software engineers to improve developer velocity, ensure production reliability, and support complex multi-environment deployments including GovCloud and air-gapped systems.

San DiegoLast seen 11 days ago
$135,000 – $180,000 · Posted 23 days ago

Entarian seeks a DevSecOps Engineer to design and maintain a DevOps Platform supporting the U.S. Navy's software development and systems integration.… The role involves developing GitLab CI/CD pipelines, building Infrastructure-as-Code with Terraform on AWS and similar cloud platforms, automating Kubernetes configurations, and implementing security controls per DoD and NIST standards. Responsibilities include hardening base images, validating Information Assurance Controls, documenting processes, and collaborating across Agile teams to deliver secure, cloud-native infrastructure.

San DiegoLast seen 21 days ago
Posted 23 days ago

Full Stack Software Engineer I at MedImpact Healthcare Systems will design, develop, and maintain software solutions across the full development lifecycle using object-oriented design principles and modern web frameworks. The role requires hands-on experience with Python, Java, microservices architecture, and web technologies (Angular/React, Spring Boot, FastAPI/Flask), along with backend infrastructure (PostgreSQL, Redis, Kafka, Kubernetes).… You'll work in an Agile Scrum environment, collaborating with technical and non-technical teams, and communicate design decisions to diverse audiences. The position is both client-facing and internal-focused, requiring strong communication skills and the ability to work independently and collaboratively in a fast-paced healthcare technology environment.

San DiegoLast seen 1 day ago
$150,000 – $165,000 · Posted 27 days ago

Design, build, and maintain secure enterprise data platforms supporting AI/ML workloads, with responsibility for database optimization, CI/CD automation, and DoD compliance. Mentor junior engineers and provide technical leadership on architecture decisions.… Implement security controls aligned with NIST 800-53, RMF, and DISA STIG standards throughout the platform lifecycle. Develop ETL/ELT pipelines, support data lakes and lakehouses, and evaluate emerging data technologies for mission fit.

San DiegoLast seen 25 days ago
Posted 27 days ago

Design and build platform engineering infrastructure, tools, and automation for cloud-native environments on AWS, using Infrastructure as Code (Terraform/CDK), CI/CD, and observability practices. Develop reusable modules, APIs, CLIs, and service templates that improve developer experience and enforce security, compliance, and cost controls.… Establish monitoring, logging, alerting, SLOs, and incident response practices; participate in 24x7 on-call rotation with a focus on production reliability and automation. Requires 7+ years in DevOps, SRE, platform engineering, or cloud operations with deep hands-on experience in multi-account AWS setups, GitOps, and observability platforms like Datadog or Prometheus.

San DiegoLast seen 25 days ago
Posted 28 days ago

The Site Reliability Engineer owns the reliability, scalability, and operational readiness of production services running on AWS and Kubernetes. Responsibilities include designing highly available infrastructure with Terraform and managed AWS services, building CI/CD pipelines with GitHub Actions and Argo CD, defining SLOs and implementing observability with Prometheus and Grafana, leading incident response, and automating operational work with Python, Go, or Bash.… The role requires 3–8 years of hands-on SRE or DevOps experience, production Kubernetes expertise, strong AWS and infrastructure-as-code knowledge, and proficiency with observability and deployment strategies.

San DiegoLast seen 26 days ago
$154,000 – $231,000 · Posted 1 month ago

Design, build, and operate an enterprise Kubernetes platform at scale across Qualcomm, establishing organizational standards for cluster operations, GitOps workflows with Argo CD, and security policies using Kyverno and Cilium. Own platform reliability, observability (Datadog, Prometheus/Grafana), and cost efficiency while serving as technical authority for security governance and multi-cloud strategy.… Lead cross-organizational teams without direct authority, mentor staff engineers, drive technology evaluation across the CNCF landscape, and participate in on-call incident response including 24/7 coverage.

San DiegoLast seen 1 day ago