Unknown Company

Site Reliability Engineer, GNC

hawthorne, ca • Posted 3 weeks ago
Onsite Full Time Space Travel

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER, GNCSpaceX’s mission is to make humanity multiplanetary by developing fully and rapidly reusable launch systems capable of launching Starship multiple times per day while continuing to scale the Starlink constellation. To support these goals, we are seeking a Site Reliability Engineer to operate and scale custom-built, mission-critical products for the Guidance, Navigation, and Control (GNC) teams.GNC teams at SpaceX are responsible for vehicle design, trajectory design and optimization, high-fidelity vehicle simulation, software and control algorithm development, while also supporting both launch and on-orbit operations across multiple vehicle programs. In this role, you will work closely with GNC teams across SpaceX to maintain and improve a suite of critical GNC-focused tools and infrastructure that must scale reliably to enable a multiplanetary future. These systems include on-prem services, large-scale Monte Carlo simulations on our high-performance computing (HPC) cluster, automated data analysis pipelines, continuous integration systems for rocket and simulation software, GNC analysis infrastructure, and vehicle configuration verification tools.The ideal candidate is flexible, possesses broad skills spanning product operations and software development, and thrives in a fast-paced, high-impact environment.RESPONSIBILITIES:Deploy, upgrade, operate, and scale a suite of mission-critical GNC products and servicesProvision and maintain virtual and physical serversWork with SpaceX HPC team to monitor and maintain an HPC cluster consisting of tens of thousands of CPUs.Closely collaborate with GNC software engineers to create highly operable and maintainable productsMonitoring and incident response for web applications and servicesManage the underlying computational infrastructure of GNC in collaboration with IT stakeholdersEngage in and improve the whole lifecycle of services from whiteboard to operationalMake data-driven recommendations for future hardware purchasesPractice sustainable incident response and postmortemsProvide end-user support to GNC engineering for products by becoming an expert on analysis applications and support users in troubleshooting and pointing to featuresConfigure automated deployment pipelines for web appsDevelop or improve GNC web apps and tools for better usability, maintainability, and robustnessDemo and document new software changes such as operating system upgrades, shared filesystem changes, or major tool rolloutsFocus on performance bottlenecks and performance improvement techniquesBASIC QUALIFICATIONS:Bachelor’s degree in computer science, information systems/IT, engineering, math, or scientific discipline and 2+ years of software development experience OR 4+ years of professional experience building software with site reliability or DevOps in lieu of a degree1+ years of experience with Linux operating systems1+ years of experience with Python and Python based development frameworksPREFERRED SKILLS AND EXPERIENCE:2+ years of systems administration, site reliability engineering, or DevOps experience2+ years of experience with Python and Python-based development frameworks2+ years of Linux experienceExpertise with Docker, Vagrant, and Kubernetes or similar technologiesExtensive Experience with configuration management tools such as Ansible, Puppet, TerraformExperience with build systems (Make, Bazel / Pants / Buck, Gradle) and package management tools (pip, npm)Strong understanding of virtualization and hypervisor technologiesUnderstanding of databases and data modelingExperience with automatically managing dozens or hundreds of serversStrong networking knowledge of TCP/IPExperience scaling web applications and optimizing applications for performanceExperience with managing on-prem infrastructure, including direct experience managing GPU fleetsExperience with high-performance computing systems or large-scale data analysis systemsMust be comfortable working with mission-critical and sensitive systems, with a sense of urgency appropriate to the responsibilitiesAbility and willingness to obtain a Top Secret clearanceADDITIONAL REQUIREMENTS:An active clearance may provide the opportunity for you to work on sensitive SpaceX missions; if so, you will be subject to pre-employment drug and random drug and alcohol testing Willing to work extended hours and weekends when needed to meet critical deadlinesCOMPENSATION AND BENEFITS:Pay Range:Site Reliability Engineer/Level I: $125,000.00 - $145,000.00/per yearSite Reliability Engineer/Level II: $145,000.00 - $175,000.00/per yearYour actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.Base salary is just one part of your total rewards package at SpaceX.

You may also be eligible for long-term incentives, in the form of company stock, stock options, or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.ITAR REQUIREMENTS:To conform to U.S. Government export regulations, applicant must be a (i) U.S.

citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State.

Learn more about the ITAR here.  SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to   

Site Reliability Engineer, GNC in hawthorne at Unknown Company

This position is listed as full time and onsite.

Back to Job Search