ByteDance is seeking engineers for its Inference Infrastructure team to advance the scale and efficiency of AI inference across multi-cloud data centers. You will work on Kubernetes-native control planes, GPU-accelerated ML platforms, and open-source tooling to enable production-grade AI workloads.
The role emphasizes collaboration across global teams, contribution to open-source, and tackling ambitious cloud-native challenges in a hyper-scale environment.
#J-18808-LjbffrGraduate Cloud-Native LLM Inference Engineer in san jose at Unknown Company
This position is listed as full time and onsite.