Wayve in London is seeking a Staff ML Performance Engineer to optimise inference for edge accelerators and GPUs, enabling our first driving product. You will help define the technical direction to run large transformer models reliably on in-vehicle compute, spanning ML systems, compilers, runtimes, kernels and embedded deployment.
You will collaborate with model developers to influence architecture and deployment choices, build benchmarking and tooling, and deliver measurable performance
#J-18808-Ljbffr