Nebius is building a high-performance inference and fine-tuning platform in a leading GPU cloud. We aim to push foundation models to hardware limits with maximum throughput, minimal latency, and optimised cost-per-token across thousands of GPUs.
Join Token Factory within Nebius Cloud to work on inference optimization, low-precision training, and novel architectures, contributing to open-source and in-house components. A strong ML/engineering background is essential.
#J-18808-Ljbffr