AMD in San Jose, CA seeks a performance-obsessed engineer to drive AI inference performance on AMD GPUs. Lead a small team, profile and optimize models end-to-end, and uplift customer engagements with measurable results.
You will diagnose kernel-level bottlenecks, optimize multi-node distributed inference, and upstream efforts to open-source frameworks while collaborating with customers and internal teams.
#J-18808-LjbffrAI Inference Performance & Scale Engineer in san jose at Unknown Company
This position is listed as full time and onsite.