Unknown Company

Senior ML Inference Engineer: High-Performance GPU Systems

san francisco, ca • Posted 4 days ago
Onsite Full Time Electrical & Energy Engineering

Acceler8 Talent is recruiting an ML Inference Engineer for a Stanford-spun AI startup in San Francisco that is building an eight-figure revenue and growth trajectory. You will design, implement, and optimize the infrastructure powering large-scale LLM workloads and real-time model serving.

Ideal candidates combine Python/C++ proficiency with distributed systems experience, PyTorch expertise, and a passion for low-latency, GPU-accelerated inference at production scale.

#J-18808-Ljbffr

Senior ML Inference Engineer: High-Performance GPU Systems in san francisco at Unknown Company

This position is listed as full time and onsite.

Back to Job Search