Micron Technology is seeking an engineer to work on AI training and inference systems, focusing on LLM execution engines, memory hierarchies, and performance optimization across data-center platforms.
The role involves end‑to‑end profiling, benchmarking, and collaboration with senior engineers and researchers. You will develop tools and methods to improve throughput, latency, and resource utilization for large-scale AI workloads.
#J-18808-Ljbffr