← Back to results

observability jobs in San Diego

Posted 25 days ago

Design and review data engineering solution architectures on AWS and Databricks, aligning technical decisions with business goals and market trends. Lead enterprise-scale data migration and modernization programs, mentoring teams and driving initiatives from proposal through delivery.… Provide hands-on expertise in Databricks, PySpark, SQL, and AWS, with strong proficiency in CI/CD pipelines, data governance, and engineering best practices. Engage with clients and internal stakeholders to build productive relationships and ensure successful project execution.

San DiegoLast seen 23 days ago
Posted 25 days ago

This role is a Data Engineering Architect responsible for designing and reviewing enterprise-scale data architectures on Databricks and AWS, leading migration and modernization programs, and resolving complex technical challenges during build and deployment phases. The architect will provide strategic insights to development teams and clients, establish data governance frameworks, monitor system performance, and contribute to proposals and RFPs.… Required expertise includes hands-on proficiency in Databricks, PySpark, SQL, AWS, CI/CD pipelines, and data architecture best practices, with proven experience in enterprise-scale programs and strong technical leadership and stakeholder management skills.

San DiegoLast seen 23 days ago
$149,800 – $262,200 · Posted 25 days ago

Staff Software Engineer on the AI Experience Framework (AIUX) team, responsible for full-stack architectural decisions in a JavaScript/component-driven environment powering ServiceNow's AI-first user interfaces. You will own code from design through delivery, architect modular reusable component systems, set framework-level technical direction for AI integration and state management, and serve as an escalation point for production issues across multiple teams.… The role requires 8+ years of OO language experience (Java, C++, C#, Go), advanced expertise in modern UI frameworks (React, Angular, Vue, Lit), relational databases, and data structures/algorithms/design patterns at scale.

San DiegoLast seen 23 days ago
$149,800 – $262,200 · Posted 26 days ago

Staff Software Engineer role owning full-stack architecture for ServiceNow's AI Experience Framework (AIUX), building modular Lit-based web components that power conversation-first applications. Responsible for end-to-end code delivery from design through production, setting technical direction for framework-level APIs, and serving as an escalation point for production issues across multiple teams.… Requires 8+ years of OO language experience (Java, C++, C#, Go), advanced JavaScript/modern UI framework expertise (Angular, React, Vue, Lit), and deep knowledge of data structures, algorithms, design patterns, and performance optimization. Expected to mentor colleagues, coordinate cross-org architecture decisions, and architect scalable component systems for extensibility.

San DiegoLast seen 23 days ago
$105,780 – $189,347 · Posted 26 days ago

The MLOps Engineer II designs, develops, and operates scalable machine learning infrastructure and deployment pipelines on AWS, working with data scientists and cloud engineers to productionize ML models. Responsibilities include building ML pipelines using SageMaker, Lambda, Step Functions, and S3; implementing CI/CD pipelines with GitHub and AWS tools; developing Infrastructure-as-Code using CloudFormation, Terraform, or AWS CDK; and ensuring production ML systems are reliable, secure, and cost-efficient.… The role requires strong Python coding skills, hands-on AWS experience, production ML deployment experience, and the ability to independently implement technical solutions. Candidates will also lead FinOps optimization, implement monitoring and alerting, ensure security and compliance, and mentor junior engineers.

San DiegoLast seen 24 days ago
Posted 26 days ago

As a Fullstack Software Engineer, you will design and build modern web applications and APIs for financial technology, working across ReactJS/TypeScript frontends and backend systems (Node.js, Java/Spring) that handle large-scale transactions with real-time processing. You'll own end-to-end features, implement secure and compliant systems with advanced security practices (CSP, encryption, vulnerability assessment), and optimize for performance and scalability.… The role requires 5+ years of software engineering experience with strong proficiency in React, modern JavaScript/TypeScript, backend frameworks, SQL/NoSQL databases, and testing frameworks (Jest, Cypress, Playwright). You'll collaborate with engineers, designers, and product managers in an agile environment, participate in code reviews, and mentor junior developers.

San DiegoLast seen 24 days ago
$129,600 – $194,400 · Posted 26 days ago

Staff IT Cloud Engineer at Qualcomm responsible for designing, implementing, and operating enterprise cloud environments across AWS, OCI, and Azure. Establish CI/CD/GitOps operating models, manage Kubernetes platforms, drive cloud governance and FinOps practices, and mentor engineers on infrastructure modernization and platform reliability.… Requires 10+ years of cloud/DevOps experience with advanced Terraform, Kubernetes, and enterprise cloud platform expertise; strong hands-on background in AWS, OCI, or Azure with deep understanding of cloud networking, security, and observability.

San DiegoLast seen 24 days ago
$158,000 – $194,000 · Posted 26 days ago

As a Software Architect, you will design AI-first software development lifecycles and lead teams in building GenAI-integrated platforms. You'll architect solutions using prompt engineering, RAG, embeddings, and agentic AI while managing multi-developer environments where AI agents and engineers collaborate.… You'll own product architecture, prototype AI-driven solutions with modern DevOps and observability, and coach teams on AI-accelerated development practices. The role requires 6+ years of production software experience, hands-on use of agentic IDEs, proven GenAI integration delivery, and strong architecture expertise across APIs, microservices, cloud platforms, and CI/CD.

San DiegoLast seen 24 days ago
$293,200 – $439,800 · Posted 26 days ago

VP-level leadership role overseeing end-to-end software architecture and development for Qualcomm's data center CPU platforms, including server firmware, networking, platform management, and cloud-native infrastructure. Responsible for defining the software roadmap spanning compilers, runtimes, orchestration, lifecycle management, and customer deployment readiness.… Will build and scale a global software engineering organization, establish engineering best practices and release discipline, and serve as a senior technical engagement point for hyperscaler and enterprise customers. Requires 15+ years of systems/infrastructure/data center software experience and 6+ years managing engineering teams and complex cross-functional execution.

San DiegoLast seen 24 days ago
$116,600 – $194,400 · Posted 26 days ago

As a Staff Data Engineer on Dexcom's Commercial Data Science & Revenue Operations team, you will design and operate the company's commercial data platform on AWS, building pipelines that ingest from 12+ vendor sources, transform data, and deliver analytics-ready datasets to data scientists, analysts, and 300+ field representatives. You will own AWS-native development using Lambda, Glue, CDK, Step Functions, and Redshift, with responsibility for data quality frameworks, alerting systems, warehouse schema design, and governance compliance.… The role requires hands-on proficiency in Python, SQL, AWS architectures, and CI/CD practices, along with experience building observability infrastructure and working with healthcare compliance requirements like HIPAA. You will also mentor team members and partner with business stakeholders to translate data capabilities into strategy.

San DiegoLast seen 24 days ago
Posted 26 days ago

Design and build platform engineering infrastructure, tools, and automation for cloud-native environments on AWS, using Infrastructure as Code (Terraform/CDK), CI/CD, and observability practices. Develop reusable modules, APIs, CLIs, and service templates that improve developer experience and enforce security, compliance, and cost controls.… Establish monitoring, logging, alerting, SLOs, and incident response practices; participate in 24x7 on-call rotation with a focus on production reliability and automation. Requires 7+ years in DevOps, SRE, platform engineering, or cloud operations with deep hands-on experience in multi-account AWS setups, GitOps, and observability platforms like Datadog or Prometheus.

San DiegoLast seen 24 days ago
$141,000 – $212,000 · Posted 27 days ago

Design, build, and operate scalable platform infrastructure across Azure, AWS, and private cloud environments using infrastructure-as-code, Kubernetes, and automation tooling. Develop reusable platform capabilities, CI/CD pipelines, and self-service infrastructure for engineering teams.… Own platform initiatives end-to-end from technical design through production operations, including deployment automation, configuration management, observability, and lifecycle management. Requires 7+ years of platform engineering or DevOps experience, strong hands-on proficiency with Terraform, Ansible, Python/Go, Kubernetes, and Linux systems administration.

San DiegoLast seen 25 days ago
Posted 27 days ago

The Site Reliability Engineer owns the reliability, scalability, and operational readiness of production services running on AWS and Kubernetes. Responsibilities include designing highly available infrastructure with Terraform and managed AWS services, building CI/CD pipelines with GitHub Actions and Argo CD, defining SLOs and implementing observability with Prometheus and Grafana, leading incident response, and automating operational work with Python, Go, or Bash.… The role requires 3–8 years of hands-on SRE or DevOps experience, production Kubernetes expertise, strong AWS and infrastructure-as-code knowledge, and proficiency with observability and deployment strategies.

San DiegoLast seen 25 days ago
$130,000 – $150,000 · Posted 27 days ago

Lead the modernization of legacy SSIS/SSRS data systems into cloud-native Lakehouse pipelines using Databricks, dbt, Fivetran, and Airflow. Design standardized frameworks for ingestion, transformation, and consumption layers while driving CI/CD, observability, and engineering best practices across Azure and GCP.… Evaluate emerging technologies (Delta Live Tables, Iceberg, streaming ingestion) through POCs and POVs, and leverage AI-assisted tools (Databricks Assistant, Cursor AI, GitHub Copilot) to accelerate development and reduce technical debt. Provide technical guidance to the team and translate business requirements into scalable data solutions.

San DiegoLast seen 25 days ago
$171,900 – $300,800 · Posted 28 days ago

Lead a team of engineering managers and software engineers building distributed agentic systems and AI-powered developer platforms at ServiceNow. You'll drive technical strategy for highly scalable, fault-tolerant systems; mentor teams; and architect foundational capabilities that improve developer experience for AI-first development.… This role requires 8+ years of software engineering experience with proven distributed systems expertise, 3+ years managing high-performing teams, and deep hands-on knowledge of Node.js, TypeScript, Kubernetes, and cloud-native infrastructure.

San DiegoLast seen 25 days ago
$135,375 – $135,375 · Posted 28 days ago

Principal-level individual contributor role leading AI/ML architecture and engineering strategy at scale. Requires 13+ years across ML, data engineering, and distributed systems, with 3+ years shipping production Generative/Agentic AI systems.… Hands-on expertise in RAG, vector databases, LLM optimization, agentic orchestration, and MLOps/LLMOps practice. Will set technical direction, own architecture decisions, drive platform initiatives, and raise engineering standards across the organization without direct people-management responsibility.

San DiegoLast seen 5 days ago
Posted 1 month ago

As a Staff AI Engineer, you will lead the design and evolution of Teradata's enterprise AI platforms, owning architecture decisions for agentic AI, LLM systems, RAG pipelines, and vector stores. You will operate with high autonomy, partner with product and architecture leadership, and drive technical strategy across multiple engineering teams.… The role requires 8+ years building backend services, distributed systems, or data/AI platforms, with strong proficiency in Java, Go, or Python and deep expertise in distributed system design, cloud-native architectures, and production AI systems. You will also mentor senior engineers, establish engineering standards, and serve as a technical escalation point for complex system design and reliability challenges.

San DiegoLast seen 28 days ago
$187,000 – $187,000 · Posted 1 month ago

Lead a data engineering team responsible for building scalable data pipelines, trusted datasets, and reusable data products powering analytics, experimentation, and AI across Scribd. This role blends technical leadership with hands-on architecture guidance; you'll establish engineering standards, drive design reviews, mentor engineers, and partner cross-functionally to translate business needs into production-grade data solutions.… Requires 10+ years in data engineering or data platforms, 3+ years leading engineering teams, deep expertise in dimensional modeling and scalable data architectures, and strong SQL plus Python/Scala skills. You'll work with modern cloud data platforms like Databricks, Snowflake, and BigQuery, and distributed processing frameworks like Spark.

San DiegoLast seen 29 days ago
$180,000 – $254,000 · Posted 1 month ago

Director-level role leading a software engineering team at a life sciences company building multi-omic analysis platforms. Requires hands-on full-stack expertise (React, Go, AWS, Temporal) alongside team management, architecture ownership, and end-to-end product delivery.… Must drive regulated software compliance (IEC 62304, ISO 13485, FDA 21 CFR Part 820), design controls, and V&V documentation while prototyping with AI tools and scaling agentic workflows within the team.

San DiegoLast seen 8 days ago
$160,000 – $220,000 · Posted 1 month ago

Senior Software Engineer responsible for designing and implementing scalable solutions for a proprietary learning management system (LMS), modernizing application architecture, and building high-quality full-stack software. The role requires 8+ years of professional software development experience and 5+ years designing enterprise-scale applications, with deep expertise in .NET, C#, Azure, React, TypeScript, and microservices architecture.… Responsibilities include leading architectural decisions, supporting CI/CD and automation initiatives, participating in code reviews, reducing technical debt, and mentoring junior engineers while collaborating across Product, QA, and DevOps teams.

San DiegoLast seen 29 days ago
$160,000 – $220,000 · Posted 1 month ago

RIVO is seeking a Senior Data Platform Engineer to architect and build a modern enterprise data platform supporting analytics, machine learning, and near real-time processing. You will design the foundational data warehouse or lakehouse architecture, modernize legacy ETL processes, implement CDC-based ingestion and low-latency pipelines, and establish data quality and governance frameworks.… This role requires 8+ years of data warehousing or platform engineering experience, expertise in cloud platforms (Snowflake, Databricks, Azure), SQL and Python proficiency, and the ability to support both analytics and data science workloads at enterprise scale.

San DiegoLast seen 29 days ago
$180,000 – $254,000 · Posted 1 month ago

Director of Software Engineering leading end-to-end delivery across customer-facing products and internal tooling at a life-sciences company. This role combines hands-on full-stack development (React, Go, AWS, Temporal) with team leadership, architecture ownership, and regulated software compliance (IEC 62304, ISO 13485, FDA 21 CFR Part 820).… Candidates should have 10+ years of software engineering with 3+ years managing teams, fluency with modern AI coding assistants, and experience shipping instrumentation, cloud SaaS, and ML/AI features.

San DiegoLast seen 1 month ago
Posted 1 month ago

A Forward Deployed AI Engineer who builds and deploys production AI/ML and Generative AI applications, particularly agentic workflows and RAG systems integrated with enterprise data and business systems. The role requires 6+ years of software/ML/AI engineering experience, advanced Python proficiency, hands-on expertise with LLMs, modern GenAI frameworks (LangChain, LangGraph, LlamaIndex), and cloud platforms (Azure, AWS, GCP).… You'll own solutions from prototype through production, working directly with product and business stakeholders to identify where AI creates value, troubleshoot integration issues, and drive continuous improvement.

San DiegoLast seen 11 days ago
$116,600 – $194,400 · Posted 1 month ago

Senior software engineer focused on building cloud-based backend services and data platforms for medical device systems at Dexcom. Design and develop end-to-end data pipelines supporting large-scale ingestion, transformation, and storage, while building and maintaining APIs that expose data to internal systems, partners, and mobile apps.… Collaborate with Data Engineers, AI/ML Engineers, and Platform teams to integrate AI insights into production systems, ensuring high standards for reliability, performance, scalability, and data quality. Lead through code reviews, design mentorship, and technical guidance while contributing to architecture decisions and long-term platform roadmap.

San DiegoLast seen 1 month ago
$190,000 – $280,000 · Posted 1 month ago

Shield AI seeks an experienced Site Reliability Engineer to establish and mature the SRE function across Hivemind's cloud infrastructure and platform services. This hands-on leadership role involves defining reliability targets (SLIs/SLOs), improving observability and incident response practices, investigating complex failures, and driving adoption of reliability-first engineering practices.… The SRE Lead will mentor teams, develop operational automation tooling, and manage the multi-quarter SRE roadmap while working closely with Cloud Engineering and product teams to enhance system resilience and operational excellence.

San DiegoLast seen 1 month ago