Lever, Inc. is seeking an AI Research Engineer (Kernel & Inference Optimization) in Spain to push the boundaries of efficient AI systems.
You will work at the intersection of AI research, systems engineering, and high-performance model inference, delivering practical improvements across diverse hardware environments. You will develop and optimize model-serving architectures, write custom GPU kernels for mobile hardware (MSL), and apply pruning, quantization, and Flash Attention to reduce latency
#J-18808-LjbffrEdge AI Inference Architect: Kernel & Performance in workfromhome at Unknown Company
This position is listed as full time and onsite.