← Back to results

etl jobs in San Diego

$85,000 – $100,000 · Posted 1 month ago

Design, build, and maintain Python and SQL-based data transformation workflows using Databricks and PySpark for large-scale processing. Analyze datasets to independently identify actionable insights, translate ambiguous data problems into clear analytic solutions, and ensure data quality and governance across pipelines.… Collaborate with stakeholders and cross-functional teams to integrate data outputs into downstream applications, troubleshoot pipeline issues, and document architecture and processes.

San DiegoLast seen 29 days ago
$85,000 – $100,000 · Posted 1 month ago

Design and build Python and SQL-based data transformation workflows using Databricks and PySpark for large-scale processing. Analyze unstructured datasets to independently identify actionable insights, translate ambiguous data problems into clear analytic solutions, and ensure data quality and governance across pipelines.… Troubleshoot pipeline issues, document processes and architecture, and partner with cross-functional teams to integrate data outputs into downstream applications and dashboards.

San DiegoLast seen 1 month ago
$125,000 – $145,000 · Posted 1 month ago

Design and build production ETL/ELT pipelines ingesting data from databases, APIs, applications, and Microsoft 365 sources into Azure Databricks and Power BI. Own code quality, CI/CD automation, testing strategies, monitoring, and data validation across development, test, and production environments in Azure DevOps.… Develop transformations and curated data products for analytics and AI applications, including secure document ingestion, metadata indexing, and telemetry. Collaborate with business and technical stakeholders to communicate progress, data limitations, and risks while maintaining runbooks and technical documentation.

San DiegoLast seen 1 month ago
$110,000 – $162,000 · Posted 1 month ago

Design and build agentic systems that automatically acquire, clean, format, and quality-control large-scale biomedical datasets for training a multimodal transformer model. You will develop LLM-based agents that generate trustworthy pipeline code, implement automated quality-control workflows, and maintain distributed data pipelines in production.… The role requires strong Python engineering, hands-on experience with LLM APIs and agentic patterns, familiarity with biomedical data formats, and data engineering fundamentals including ETL design and validation at scale.

San DiegoLast seen 1 month ago
$110,000 – $162,000 · Posted 1 month ago

Design and build agentic systems that automatically acquire, clean, format, and quality-control large-scale biomedical datasets for the Enchant multimodal transformer model. You will develop LLM-based agents that generate reproducible data pipeline code, implement automated quality-control workflows, and maintain distributed pipelines in production.… The role combines strong software engineering with scientific understanding of biomedical data formats and ETL design, working at the intersection of LLM automation and data engineering.

San DiegoLast seen 1 month ago
$116,600 – $194,400 · Posted 1 month ago

Senior software engineer focused on building cloud-based backend services and data platforms for medical device systems at Dexcom. Design and develop end-to-end data pipelines supporting large-scale ingestion, transformation, and storage, while building and maintaining APIs that expose data to internal systems, partners, and mobile apps.… Collaborate with Data Engineers, AI/ML Engineers, and Platform teams to integrate AI insights into production systems, ensuring high standards for reliability, performance, scalability, and data quality. Lead through code reviews, design mentorship, and technical guidance while contributing to architecture decisions and long-term platform roadmap.

San DiegoLast seen 1 month ago
$110,000 – $162,000 · Posted 1 month ago

Design and build agentic systems that automatically acquire, clean, format, and quality-control large-scale biomedical datasets for training the Enchant multimodal transformer model. You will develop LLM-based agents that generate trustworthy pipeline code, implement automated QC workflows, and maintain distributed data pipelines in production.… The role requires strong Python engineering, hands-on experience with LLM APIs and agent patterns, familiarity with biomedical data formats, and data engineering fundamentals including ETL design and validation at scale.

San DiegoLast seen 1 month ago
$130,813 – $130,813 · Posted 1 month ago

This role designs and governs enterprise data architectures across SQL, NoSQL, and cloud platforms for a financial institution. The Data Architect will establish data standards, lead golden-record and SSOT initiatives, architect ETL/ELT pipelines, manage data quality and governance frameworks, and oversee database platform strategy (SQL Server, Azure SQL, Snowflake, MongoDB Atlas, Cosmos DB).… The position requires 12+ years of progressive experience, deep cloud data platform expertise, SQL and Python proficiency, and demonstrated leadership guiding engineering teams through enterprise modernization and compliance-driven initiatives.

San DiegoLast seen 1 month ago
$132,000 – $132,000 · Posted 1 month ago

Senior Enterprise Data Engineer role requiring 5+ years designing and developing OLTP/OLAP databases, ETL/ELT pipelines using SSIS, Azure Data Factory, or Python, and enterprise reporting solutions with Power BI and SSRS. Must have advanced SQL Server expertise, strong data warehousing and dimensional modeling knowledge, and experience in Agile environments.… Preferred qualifications include Azure data platforms (Synapse, Data Lake, Fabric), Databricks, Apache Spark, Azure DevOps, and CI/CD practices.

San DiegoLast seen 1 month ago
Posted 1 month ago

Build and test agent behavior for an AI-driven service, including test case design and evaluation checks distinct from traditional software testing. Support data ingestion pipelines using AWS and BigData tools (PySpark, Airflow), collaborate with product and senior engineers to ship features, and contribute to production support (monitoring, alerting, incident response).… Requires 1–2 years of software engineering experience, a Master's in Computer Science or related field, proficiency in Python or another backend language, SQL/NoSQL fundamentals, basic REST API design, and hands-on experience with at least one LLM project (tool calling, chaining, retrieval, or multi-step coordination). Self-directed learner excited about AI tooling and focused on shipping end-to-end products.

San DiegoLast seen 1 month ago
$174,000 – $185,000 · Posted 1 month ago

Build and optimize a high-throughput compute stack for reliable, production-grade data ingestion, processing, storage, and delivery. The role focuses on making stateful, multi-threaded pipelines fast and observable under real-world constraints—profiling bottlenecks, hardening recovery behavior, and productizing machine-learning models (neural networks, tree-based, unsupervised) to meet performance, reliability, and quality standards.… You will work alongside algorithm and infrastructure engineers, using measurements to drive performance improvements and ensure production interfaces remain durable as the system evolves. Requires strong systems-level experience with concurrency, I/O, and production debugging in C++, Rust, CUDA, or C; a track record of shipping production software; and hands-on expertise in resilient data pipelines and GPU computing.

San DiegoLast seen 1 month ago
$150,000 – $180,000 · Posted 1 month ago

Senior Enterprise Applications Developer responsible for designing, delivering, and supporting complex enterprise software, integrations, data solutions, and automation across cloud and on-premises environments. Requires 7+ years of progressive experience with advanced Python proficiency, hands-on expertise in REST APIs, ETL/data pipelines, CI/CD practices, and AWS or Azure cloud-native development.… Must demonstrate ability to lead solution design from requirements through production support, with experience in workflow automation, AI/LLM integrations, and compliance frameworks (HIPAA, SOC 2, SOX). Strong emphasis on security, governance, architecture, and cross-functional collaboration with technical teams and business leaders.

San DiegoLast seen 1 month ago
$95,000 – $120,000 · Posted 1 month ago

Enterprise Applications Developer responsible for building and supporting integrations among Salesforce, SaaS platforms, cloud services, and external partners using REST APIs, web services, and event-driven patterns. Develop Salesforce enhancements using Apex, Lightning Web Components, Flow, and SOQL/SOSL.… Design and maintain data extraction, transformation, validation, and loading processes; build workflow automations; and assist with AI-enabled solutions. Troubleshoot application and integration issues, maintain technical documentation, and support compliance requirements (HIPAA, SOC 2 Type II, SOX).

San DiegoLast seen 1 month ago
$189,200 – $372,900 · Posted 1 month ago

Lead a team of Databricks engineers delivering GenAI and LLM solutions to government clients. You'll architect and deploy production-scale data platforms, mentor engineers, and translate business requirements into AI solutions using Databricks' full stack (Lakehouse, Agent Bricks, Model Serving, Genie, Apps).… Enforce strong data management, CI/CD, testing, and documentation practices while working across AWS, Azure, and Google Cloud environments.

San DiegoLast seen 1 month ago
$124,000 – $190,000 · Posted 1 month ago

Design, build, and maintain scalable data pipelines and platforms that power high-volume, data-driven applications. Develop and optimize ETL/ELT processes for large, complex datasets using Databricks and Spark, and build backend data services in C# / .NET.… Lead technical design decisions, mentor engineers, and leverage AI-assisted coding tools to improve engineering efficiency. Requires 5+ years of data engineering experience, strong expertise with data pipelines and ETL frameworks, and hands-on experience with AWS-based data platforms.

San DiegoLast seen 14 days ago
$80,000 – $90,000 · Posted 1 month ago

A junior data engineer will contribute to data pipeline development across AWS, Azure, Salesforce, and MuleSoft, learning the full engineering lifecycle with guidance. Core responsibilities include building and maintaining data pipelines, implementing Kimball-based data models, writing Python orchestration and data-quality scripts, and supporting Power BI analytics.… The role emphasizes AI-assisted development using Claude Code, Cursor, and GitHub Copilot, with requirements for 0–2 years of experience, 6+ months of Python, intermediate SQL, at least one end-to-end pipeline project, and Power BI/Tableau dashboard experience.

San DiegoLast seen 1 month ago
$141,600 – $212,400 · Posted 1 month ago

Staff Data Engineer at Illumina seeking a senior individual contributor with 10+ years building and scaling data products on Databricks and Snowflake. The role spans end-to-end ownership of data pipelines across Supply Chain, Manufacturing, and Quality domains, requiring advanced Python, SQL, and data modeling skills with hands-on expertise in Spark, Delta Lake, dbt, and distributed systems design.… Responsibilities include designing medallion-architecture data products, embedding governance and data quality, adopting AI in analytics workflows, and mentoring a global engineering team. The candidate will set technical standards, lead architecture decisions, and communicate trade-offs across business, AI, and platform stakeholders.

San DiegoLast seen 14 days ago
$95,378 – $160,850 · Posted 1 month ago

This Data Engineer II role focuses on data operations and infrastructure, supporting the transition of data systems from development to production. The engineer will administer data lakehouse/warehouse platforms (Snowflake, Databricks, Redshift, BigQuery, Azure Synapse), manage RBAC and user governance, monitor and troubleshoot data pipelines, and contribute to data governance and quality best practices.… Key responsibilities include orchestration tool administration (Apache Airflow, dbt), observability framework development, incident management, and hands-on SQL and Python work to maintain and optimize production data systems.

San DiegoLast seen 15 days ago
$175,000 – $185,000 · Posted 1 month ago

Principal Engineer leading the design, build, and operation of an enterprise data platform serving 50+ source systems across R&D, Commercial, Manufacturing, and other business functions. The role combines hands-on data engineering (PySpark, T-SQL, Python, Data Factory, medallion architecture) with technical leadership of data engineers, vendor management, platform standards enforcement, and cross-functional stakeholder partnership.… Responsibilities span CI/CD pipeline design, data quality and reliability instrumentation, semantic layer development, data governance implementation, and delivery of new source-system integrations at scale.

San DiegoLast seen 1 month ago
$150,000 – $190,000 · Posted 1 month ago

Clinical Data Engineer II/III will develop and maintain data transformation models across clinical schemas, ensuring data quality, reliability, and regulatory compliance in a healthcare environment. The role requires designing data contracts, building reconciliation logic to unify multi-source data, and owning clinical result data end-to-end from ingestion through monitoring.… You will partner with clinical operations, regulatory, and research teams to resolve data issues, maintain dashboards subject to formal validation, and serve as a subject-matter expert on the warehouse clinical data. This position demands 5+ years of hands-on data engineering with production transformation frameworks (dbt preferred), deep SQL expertise, and demonstrated experience in regulated environments.

San DiegoLast seen 1 month ago
Posted 1 month ago

Senior developer responsible for designing, building, and optimizing business intelligence dashboards and reports using Tableau, SAP BusinessObjects (BOBJ), and related tools. Primary responsibilities include gathering client requirements, designing dashboards connected to multiple data sources (Oracle, SAP BW, SAP HANA), creating BEx queries and process chains, optimizing query performance, and managing security and user access.… Must have 5+ years of SAP BOBJ and Tableau experience, 3+ years with Power BI, and strong expertise across SAP design tools (Design Studio, Lumira), Crystal Reports, WebI, Spotfire, and modern web technologies (HTML5, CSS3, JSON, SAPUI5).

San DiegoLast seen 1 month ago
$158,400 – $237,600 · Posted 1 month ago

Design, build, and maintain scalable batch and streaming data pipelines on AWS and Databricks, developing reusable data engineering frameworks and standardized patterns for ingestion, transformation, and validation. Leverage AI-based techniques to automate data quality checks, anomaly detection, and pipeline self-healing.… Define SLIs/SLOs, participate in on-call rotations, and mentor junior engineers while driving adoption of data frameworks and best practices across analytics, reporting, and AI/ML teams.

San DiegoLast seen 15 days ago
$61,900 – $141,000 · Posted 1 month ago

As a Data Engineer at Booz Allen Hamilton, you will develop and deploy data pipelines and platforms that organize disparate data sources to yield actionable insights for mission-critical applications. You'll work with Python, SQL, or similar languages to build ETL operations, manage relational and non-relational databases, and support analytics workloads on cloud platforms.… The role requires 1+ years of data engineering, ETL, or pipeline development experience, proficiency with source control (GitHub/Atlassian), and Linux/Windows scripting. You'll collaborate with analysts, developers, and data scientists in an agile environment to design, develop, and maintain scalable data solutions.

San DiegoLast seen 1 month ago
$140,000 – $150,000 · Posted 1 month ago

Own the product direction, roadmap, and prioritization for a ground-up AI-accelerated migration platform that helps law firms move onto the CARET Legal practice management system. You will translate complex migration and onboarding challenges into well-scoped product opportunities, work closely with an Engineering Tech Lead on technical architecture, and partner with migration/operations teams to deliver measurable customer and business outcomes.… The role requires deep technical fluency with data flows and system behavior, strong judgment about scope and tradeoffs, and the ability to engage directly with engineers and cross-functional stakeholders. You'll define success metrics around time-to-value, cutover experience, delivery predictability, and quality while balancing platform investment with customer commitments.

San DiegoLast seen 1 month ago
$138,000 – $224,400 · Posted 1 month ago

Build and integrate oncology applications across laboratory systems, data pipelines, and enterprise platforms at Eli Lilly. Develop AI-enabled capabilities for data retrieval, extraction, and analysis while owning application reliability, user support, and lifecycle planning.… Deploy on Lilly cloud infrastructure using modern server-side frameworks (Python/Django/Flask preferred), front-end technologies (JavaScript/TypeScript), relational databases, and CI/CD pipelines. Mentor junior engineers and collaborate directly with scientific and operational stakeholders.

San DiegoLast seen 1 month ago