← Back to results

autoscaling jobs in San Diego

Posted 1 day ago

Staff-level backend engineer responsible for designing and maintaining statistical capacity models that forecast infrastructure resource needs (CPU, memory, connections, ElasticCache, DynamoDB, AuroraDB) to support traffic demands during peak sales events. Sets technical strategy for the team, collaborates across product and analytics to ensure system sustainability, owns operational excellence including monitoring and on-call processes, and establishes code review and design standards.… Requires 8+ years building highly available distributed systems at scale using Python or Kotlin, AWS, MySQL, Spark, and Kubernetes, with proven experience scaling systems for major sales events like Prime Day and Black Friday.

San DiegoLast seen today
$143,000 – $215,000 · Posted 12 days ago

Senior Platform Engineer to design, build, and operate a cloud-native compute platform on AWS and Kubernetes. You will own platform security, scalability, and evolution through strategic initiatives and migration efforts, lead incident response and troubleshooting, and provide technical leadership and mentorship.… Requires 8+ years in platform/SRE/infrastructure engineering with deep AWS expertise, production Kubernetes experience, hands-on CI/CD (GitHub, GitHub Actions, Argo CD), Infrastructure as Code (Terraform/CDK), Kubernetes networking knowledge, and production observability/monitoring skills.

San DiegoLast seen 10 days ago
Posted 1 month ago

DevOps Engineer to build, maintain, and optimize AWS cloud infrastructure and CI/CD deployment systems, working with ECS, EKS, and EC2 workloads. The role requires hands-on experience with Infrastructure as Code (Terraform), containerization (Docker/Kubernetes), monitoring tools (Grafana), and CI/CD pipelines, with growing ownership of services and systems.… You'll troubleshoot infrastructure and application issues, collaborate with engineering teams on deployment workflows, and apply AI tooling to improve infrastructure automation and operational efficiency. The position reports to a DevOps Manager and is based in San Diego on a hybrid schedule.

San DiegoLast seen 16 days ago
$158,400 – $237,600 · Posted 1 month ago

Build and optimize scalable LLM inference platforms at Qualcomm's Cloud AI team, implementing advanced serving techniques like KV-cache management, speculative algorithms, and model optimization. Contribute to production serving frameworks (vLLM, SGLang, Triton, TGI) and work with customers on deployment solutions.… Collaborate with compiler, firmware, and platform teams to drive efficient serving through autoscaling, load balancing, and routing. Requires deep understanding of transformer architectures, strong PyTorch and Python skills, computer architecture knowledge, and hands-on experience profiling and optimizing deep learning workloads.

San DiegoLast seen 1 day ago