SpaceXAI in Memphis is seeking a Site Reliability Engineer focused on campus reliability to design monitoring, lead blameless postmortems, and drive reliability across compute, network, storage and facilities. You will coordinate with NOC and data center operations, and own runbooks and incident response for SEV-class events at the Memphis/Southaven campus.
Ideal candidates have 5+ years in SRE or plant operations in data centers or power/industrial environments, strong scripting (Python, Bash),
#J-18808-LjbffrCampus SRE: Incident Leadership & Fleet Reliability in memphis at Unknown Company
This position is listed as full time and onsite.