Senior Engineer - Machine Learning
San DiegoLast seen 1 day ago
Summary
Design, develop, and optimize machine learning systems for production AI platforms, focusing on model development, inference optimization, and scalable ML infrastructure. Responsibilities include building ML pipelines and frameworks, optimizing model inference for latency and cost, integrating LLMs into APIs and microservices, and designing data pipelines for preprocessing and feature engineering. The role requires strong software engineering fundamentals combined with deep ML expertise, working across PyTorch/TensorFlow, model-serving systems, distributed computing, and GPU environments.
Data ScienceDevOps / InfrastructureC++DockerKubernetesMachine LearningPythonRustAgentic AIApisCI CDData PipelinesDeep LearningDistributed ComputingDynamic BatchingETLFeature EngineeringGenerative AIGpuInference OptimizationKv CacheLLM Large Language ModelsMicroservicesModel DistillationModel QuantizationModel RoutingModel ServingMulti Agent OrchestrationOnnxOpensearchPyTorchQdrantRAG Retrieval Augmented GenerationRetrieval Augmented GenerationTensorFlowTransformer ArchitecturesVector DatabasesVllm