Unknown Company

Sr. Site Reliability Engineer III

washington, dc • Posted 5 days ago
Onsite Full Time IT & Technology

As a Sr. Site Reliability Engineer (SRE) III , you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.

What you’ll do:

  • Design, deploy, and maintain mission-critical application workloads on virtualized or containerized environments (e.g., VMWare or Kubernetes), ensuring scalability, availability, and compliance with government requirements.
  • Develop and sustain automated CI/CD pipelines, monitoring, and configuration management workflows to support reliable software delivery and operational observability across development, integration, staging, and production environments.
  • Provision, configure, and maintain developer environments and toolchains to support rapid, secure, and efficient development workflows, enabling mission-aligned software delivery.
  • Identify developer friction across the software development lifecycle and implement solutions to reduce that friction and provide developer-first environments.
  • Establish and maintain a high level of customer trust and confidence through deep technical expertise, and use creativity to provide innovative solutions that fit the customer’s mission needs.

What you’ll need to succeed:

  • Active Top Secret clearance, or higher.
  • Certification meeting DoD 8140 (e.g., Security+ or higher).
  • Bachelor’s degree in Computer Science or related engineering field is preferred; relevant experience may substitute.
  • 7+ years of experience in software development, systems engineering, or operations roles with responsibility for availability, performance, and reliability of production systems.
  • Demonstrated experience blending software engineering and systems administration practices to support highly available, scalable applications.
  • Experience designing and managing monitoring, alerting, and observability solutions to meet defined Service Level Objectives.
  • Experience leading or participating in incident response, root cause analysis, and continuous improvement activities.
  • Experience with Ansible and Desired State Configuration.
  • Experience with GitLab CI/CD automation and Bash scripting.
  • Experience with Kubernetes, supporting container-native storage and object storage solutions (e.g., MinIO, S3-compatible services, PortWorx).
  • Experience with enterprise load-balancing solutions (e.g., F5 or similar platforms).
  • Ability to contribute immediately with minimal ramp-up in a mission-critical operational environment.

This position is designated as essential personnel supporting continuity of operations and may require work during government shutdowns, emergencies, or other critical situations.

SALARY RANGE: $185,000 - $230,000

The salary range for this position is determined based on qualifications, skills, and relevant experience.

#J-18808-Ljbffr
Back to Job Search