The Senior Site Reliability Engineer at Nebius will focus on ensuring the reliability and performance of the inference platform, which is critical for scaling operations effectively in the AI cloud space. Responsibilities include designing telemetry pipelines, tuning Kubernetes for efficiency, and creating resilient Terraform modules. The role requires strong technical skills in Kubernetes and experience with GPU workloads, alongside the ability to collaborate and troubleshoot effectively.