Senior Site Reliability Engineer role at Nebius focused on building and maintaining the reliability, performance, and observability of the Token Factory inference platform. The ideal candidate will design telemetry pipelines, optimize Kubernetes infrastructure, and drive incident response automation for a large-scale GPU cloud serving foundation models.