The Senior Site Reliability Engineer role at Nebius focuses on enhancing the reliability, performance, and observability of the inference platform. Responsibilities include designing telemetry pipelines, tuning Kubernetes for GPU efficiency, and creating resilient infrastructure using Terraform. An ideal candidate should possess deep expertise in relevant technologies such as Kubernetes and experience with GPU workloads.