Nebius seeks a Senior ML Engineer for their Token Factory team to optimize high-performance inference and fine-tuning on their massive GPU cloud. The ideal candidate will have deep expertise in LLM optimization, transformer architectures, and inference engines to maximize throughput and minimize latency across tens of thousands of GPUs.