← Back to results

data-curation jobs in San Diego

$180,000 – $300,000 · Posted 1 day ago

As an Autonomy Engineer focused on VLA (Vision-Language-Action) pre-training, you will develop and train deep learning policies for humanoid robots, owning the full pipeline from data collection and curation to model deployment. You'll pre-train base models on diverse multi-embodiment trajectory corpora, fine-tune policies for specific tasks, work with MLOps teams to scale distributed training, and build continuous pipelines that ingest synthetic data and teleop logs.… This is primarily a deep learning role requiring 3+ years of experience shipping neural network models, hands-on expertise with LLMs, VLMs, or generative models, and strong proficiency in Python and PyTorch or JAX.

San DiegoLast seen today
Posted 1 day ago

This role supports neuroimaging research by performing manual image review and quality assurance on MRI data, conducting statistical analyses using MATLAB and Python, and maintaining data integrity across multiple acquisition sites. The analyst will design and implement data analysis tools, configure pipelines, write shell scripts and standard operating procedures, and troubleshoot data quality issues in coordination with partner institutions.… The position requires proficiency in Linux, shell scripting, MATLAB, and Python, along with hands-on experience analyzing structural and functional MRI neuroanatomic data and applying bioinformatics QC methods.

San DiegoLast seen today
Posted 1 month ago

Build and operate the data and ML infrastructure powering an AI platform for materials science, owning both sides: data pipelines that ingest and curate large-scale scientific output into training-ready formats, and model packaging, serving, monitoring, and CI/CD systems that move models safely from research to production across customer environments. You will design data ingestion and transformation workflows, implement validation and quality gates, package and version models with reproducible builds, run models through batch and online inference with safe rollout and rollback, monitor for drift and degradation, and build observability and internal tooling for engineering and science teams.… The role requires 6+ years shipping production software with deep expertise in data systems, ML infrastructure, containers, orchestration, and observability.

San DiegoLast seen 11 days ago