Unknown Company

ML Engineer—LLM Inference & GPU Optimization (Equity)

san francisco, ca • Posted 6 days ago
Onsite Full Time Electrical & Energy Engineering

IC Resources seeks an engineer to accelerate production AI systems, focusing on speed and efficiency of large language model inference. You will optimize GPU-heavy pipelines and scale distributed GPU environments in a fast-moving, innovative AI company.

You will work with research and infrastructure teams to translate cutting-edge models into reliable production systems, improving latency and throughput while leveraging modern hardware. #J-18808-Ljbffr

Back to Job Search