River AI Inc. is seeking exceptional systems engineers to build the distributed training engines behind the River API, making fine-tuning and reinforcement learning fast, numerically correct, and reliable across large GPU clusters.
You will own training workload execution, including gradient computation, optimizer updates, rollout coordination, and checkpoint recovery. You will collaborate with researchers and inference engineers to bring new learning methods into production and improve how
#J-18808-LjbffrDistributed Training Systems Engineer in palo alto at Unknown Company
This position is listed as full time and onsite.