Unknown Company

Senior ML Engineer - GPU Inference & Large-Scale AI Cloud

Remote • Posted 4 days ago
Onsite Full Time IT & Technology

Nebius is building a high-performance inference and fine-tuning platform in a leading GPU cloud. We aim to push foundation models to hardware limits with maximum throughput, minimal latency, and optimised cost-per-token across thousands of GPUs.

Join Token Factory within Nebius Cloud to work on inference optimization, low-precision training, and novel architectures, contributing to open-source and in-house components. A strong ML/engineering background is essential.

#J-18808-Ljbffr
Back to Job Search