← Back to results

Jobs at SHEIN U.S.

$92,400 – $148,800 · Posted 1 day ago

A Senior Site Reliability Engineer at SHEIN will own and operate mission-critical, large-scale distributed systems (Kubernetes, Kafka, Elasticsearch, Redis, APISIX, Nginx) running 24/7/365, participating in on-call rotations and driving fast incident response using AI-assisted log analysis and anomaly detection. The role requires strong software engineering expertise in Python or Go, deep Linux and networking knowledge, and hands-on experience with observability platforms (Prometheus, Grafana) and configuration management tools. Responsibilities include designing resilient monitoring and alerting infrastructure, automating operational workflows to eliminate toil, capacity planning, and collaborating with global teams to improve system reliability and performance.

San DiegoLast seen today
$86,400 – $138,000 · Posted 5 days ago

The Database Engineer will maintain, monitor, and optimize large-scale production MySQL database environments in distributed, mission-critical settings. Responsibilities include performance tuning, query and index optimization, database migrations and upgrades, capacity planning, and 24x7 on-call support. The role requires designing reliable database architectures, automating operational tasks, and collaborating with development teams on database design guidance. Candidates must have strong fundamentals in distributed database topologies, Linux administration, scripting proficiency (Go/Python/Shell), and experience with cloud platforms such as Azure.

San DiegoLast seen 3 days ago
$108,000 – $180,000 · Posted 9 days ago

Staff Site Reliability Engineer at SHEIN responsible for operating and evolving large-scale, mission-critical production systems with 24/7/365 on-call participation. Design, build, and maintain observability solutions (metrics, logs, traces, alerting) with AI-powered anomaly detection; own and operate core open-source infrastructure (APISIX, Nginx, Kubernetes, Kafka, Elasticsearch, Redis, Consul, Etcd, Zookeeper). Automate operational workflows, reduce incident frequency and MTTR, and provide technical leadership across global engineering teams. Requires strong software engineering skills, deep Linux/networking/distributed systems expertise, and passion for solving problems at scale.

San DiegoLast seen 7 days ago