More jobs:
Sr. Site Reliability Engineer III
Job in
Washington, District of Columbia, 20022, USA
Listed on 2026-08-09
Listing for:
Catapult Federal Services
Full Time
position Listed on 2026-08-09
Job specializations:
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Unix/Linux
Job Description & How to Apply Below
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.
What you’ll do:
- Design, deploy, and maintain mission-critical application workloads on virtualized or containerized environments (e.g., VMWare or Kubernetes), ensuring scalability, availability, and compliance with government requirements.
- Develop and sustain automated CI/CD pipelines, monitoring, and configuration management workflows to support reliable software delivery and operational observability across development, integration, staging, and production environments.
- Provision, configure, and maintain developer environments and tool chains to support rapid, secure, and efficient development workflows, enabling mission-aligned software delivery.
- Identify developer friction across the software development lifecycle and implement solutions to reduce that friction and provide developer-first environments.
- Establish and maintain a high level of customer trust and confidence through deep technical expertise, and use creativity to provide innovative solutions that fit the customer’s mission needs.
What you’ll need to succeed:
- Active Top Secret clearance, or higher.
- Certification meeting DoD 8140 (e.g., Security+ or higher).
- Bachelor’s degree in Computer Science or related engineering field is preferred; relevant experience may substitute.
- 7+ years of experience in software development, systems engineering, or operations roles with responsibility for availability, performance, and reliability of production systems.
- Demonstrated experience blending software engineering and systems administration practices to support highly available, scalable applications.
- Experience designing and managing monitoring, alerting, and observability solutions to meet defined Service Level Objectives.
- Experience leading or participating in incident response, root cause analysis, and continuous improvement activities.
- Experience with Ansible and Desired State Configuration.
- Experience with Git Lab CI/CD automation and Bash scripting.
- Experience with Kubernetes, supporting container-native storage and object storage solutions (e.g., MinIO, S3-compatible services, Port Worx).
- Experience with enterprise load-balancing solutions (e.g., F5 or similar platforms).
- Ability to contribute immediately with minimal ramp-up in a mission-critical operational environment.
This position is designated as essential personnel supporting continuity of operations and may require work during government shutdowns, emergencies, or other critical situations.
SALARY RANGE: $185,000 - $230,000
The salary range for this position is determined based on qualifications, skills, and relevant experience.
#J-18808-LjbffrTo View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×