Tower Research Capital seeks a gifted ML inference optimization engineer to bridge research and production. You will design and accelerate low-latency inference pipelines, aiming for microsecond latency across heterogeneous hardware in a high-performance trading environment.
As part of the Core Engineering team, you will evaluate platforms, optimize memory hierarchies, and collaborate with ML researchers, HPC and datacenter engineers to deploy scalable, reliable inference solutions that power
#J-18808-LjbffrHybrid Low-Latency ML Inference Engineer in new york at Unknown Company
This position is listed as full time and onsite.