Nebius is seeking a Senior ML Engineer for the Token Factory team to optimize high-performance inference and fine-tuning for large language models across tens of thousands of GPUs. The ideal candidate will have deep expertise in transformer architectures, inference optimization, and low-precision training/inference to maximize throughput and minimize latency at scale.