As a Sr. Site Reliability Engineer (SRE) III , you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.
What you’ll do:
- Design, deploy, and maintain mission-critical application workloads on virtualized or containerized environments (e.g., VMWare or Kubernetes), ensuring scalability, availability, and compliance with government requirements.
- Develop and sustain automated CI/CD pipelines, monitoring, and configuration management workflows to support reliable software delivery and operational observability across development, integration, staging, and production environments.
- Provision, configure, and maintain developer environments and toolchains to support rapid, secure, and efficient development workflows, enabling mission-aligned software delivery.
- Identify developer friction across the software development lifecycle and implement solutions to reduce that friction and provide developer-first environments.
- Establish and maintain a high level of customer trust and confidence through deep technical expertise, and use creativity to provide innovative solutions that fit the customer’s mission needs.
What you’ll need to succeed:
- Active Top Secret clearance, or higher.
- Certification meeting DoD 8140 (e.g., Security+ or higher).
- Bachelor’s degree in Computer Science or related engineering field is preferred; relevant experience may substitute.
- 7+ years of experience in software development, systems engineering, or operations roles with responsibility for availability, performance, and reliability of production systems.
- Demonstrated experience blending software engineering and systems administration practices to support highly available, scalable applications.
- Experience designing and managing monitoring, alerting, and observability solutions to meet defined Service Level Objectives.
- Experience leading or participating in incident response, root cause analysis, and continuous improvement activities.
- Experience with Ansible and Desired State Configuration.
- Experience with GitLab CI/CD automation and Bash scripting.
- Experience with Kubernetes, supporting container-native storage and object storage solutions (e.g., MinIO, S3-compatible services, PortWorx).
- Experience with enterprise load-balancing solutions (e.g., F5 or similar platforms).
- Ability to contribute immediately with minimal ramp-up in a mission-critical operational environment.
This position is designated as essential personnel supporting continuity of operations and may require work during government shutdowns, emergencies, or other critical situations.
SALARY RANGE: $185,000 - $230,000
The salary range for this position is determined based on qualifications, skills, and relevant experience.
#J-18808-Ljbffr