Senior ML Engineer role at Nebius focusing on optimizing LLM inference and fine-tuning at massive scale across tens of thousands of GPUs. The ideal candidate will have deep expertise in transformer architectures, inference optimization, and low-precision training to maximize throughput while minimizing latency and cost-per-token.