Unknown Company

LLM Inference Optimization Engineer

san francisco, ca • Posted 5 days ago
Onsite Full Time Electrical & Energy Engineering

GMI Cloud is building the leading inference optimization solution and the most advanced token platform in the global token market. We are hiring world-class Machine Learning Engineers to make GMI the new industry benchmark for LLM serving performance, cost efficiency, and production reliability.

This role focuses on frontier research, validation, and productionization of advanced inference optimization techniques, with close collaboration across platform and infrastructure teams to deliver

#J-18808-Ljbffr
Back to Job Search