The Senior Site Reliability Engineer at Nebius plays a crucial role in maintaining the reliability and performance of the inference platform within the AI cloud environment. This position focuses on designing telemetry pipelines, tuning Kubernetes for efficiency, and developing resilient infrastructure through Terraform. Candidates should have extensive experience with Kubernetes and GPU workloads, alongside strong scripting skills in Python or Bash.