The Senior Site Reliability Engineer at Nebius plays a crucial role in maintaining and enhancing the reliability and performance of the inference platform. This position involves designing telemetry pipelines, optimizing Kubernetes for GPU efficiency, and developing Terraform modules for resilience. The ideal candidate should have significant experience with infrastructure-as-code and GPU workloads, alongside strong scripting skills in Python or Bash.