d-Matrix Frontier Group in Santa Clara, CA is seeking end-to-end inference engineers to drive novel ideas to deployed, optimized systems across the inference stack. You will work from kernel-level optimization to distributed orchestration and high-level serving APIs, shaping the future of AI silicon usage.
Ideal candidates have deep experience with LLM inference, open-source frameworks, and heterogeneous hardware deployments, delivering POCs to customers and contributing to open-source projects.
#J-18808-LjbffrSenior LLM Inference Engineer - End-to-End, Equity in santa clara at Unknown Company
This position is listed as full time and onsite.