Senior ML Engineer role at Nebius focused on optimizing LLM inference and fine-tuning on massive GPU clusters. The ideal candidate has deep expertise in machine learning theory, transformer architectures, and production-scale inference optimization to maximize throughput and minimize latency across thousands of GPUs.