IC Resources seeks an engineer to accelerate production AI systems, focusing on speed and efficiency of large language model inference. You will optimize GPU-heavy pipelines and scale distributed GPU environments in a fast-moving, innovative AI company.
You will work with research and infrastructure teams to translate cutting-edge models into reliable production systems, improving latency and throughput while leveraging modern hardware. #J-18808-Ljbffr