Unknown Company

LLM Inference Systems Performance Engineer

austin, tx • Posted 4 days ago
Onsite Full Time Engineering

Micron Technology is seeking an engineer to work on AI training and inference systems, focusing on LLM execution engines, memory hierarchies, and performance optimization across data-center platforms.

The role involves end‑to‑end profiling, benchmarking, and collaboration with senior engineers and researchers. You will develop tools and methods to improve throughput, latency, and resource utilization for large-scale AI workloads.

#J-18808-Ljbffr
Back to Job Search