Staff Engineer who will design, build, and operate the Forge Platform—a distributed-systems foundation serving autonomy, ML Ops, simulation, and application teams. You will own architecture and technical standards for workflow orchestration, asynchronous processing, event-driven systems, and long-running service workflows while remaining hands-on in implementation, production troubleshooting, and reliability improvement.… Required: strong production experience with distributed systems, cloud-native platforms, or backend infrastructure; fluency in Go and Python; deep understanding of failure handling, consistency, fault tolerance, and state management; and ability to turn recurring infrastructure needs into reusable platform capabilities. You will work across Kubernetes, service networking, observability, and data pipelines to enable downstream teams to move faster on mission-critical systems.