Nebius is seeking a Senior ML Engineer to join the Token Factory team, focusing on high-performance LLM inference and fine-tuning optimization across tens of thousands of GPUs. The ideal candidate will have deep expertise in transformer architectures, inference optimization, and low-precision training to maximize throughput while minimizing latency and cost-per-token.