Long Ridge Partners is seeking a Machine Learning Performance Engineer (Inference) to architect ultra-low-latency inference pipelines for high-frequency trading. You will benchmark workloads across CPU, GPU, and FPGA, drive hardware architecture decisions, and deploy optimized kernels and libraries to maximize throughput and reduce latency in production.
The role focuses on end-to-end performance, including memory hierarchies, interconnects, and thermal/power constraints, with cross-functional
#J-18808-LjbffrHigh-Performance ML Inference Engineer (Hybrid) in new york at Unknown Company
This position is listed as full time and onsite.