Senior DevOps Engineer
Listed on 2026-09-27
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Cybersecurity
Overview
REI Systems’ mission is to deliver reliable, innovative technology solutions that advance Federal clients' missions and exceed their expectations. Our technologists and consultants are passionate about solving complex challenges that impact millions of lives. We take a Mindful Modernization approach in delivering our services, including application modernization, grants management, case management systems, government data analytics, and advisory services. This approach, the REI Way, ensures mission impact by aligning our clients' strategic objectives with measurable outcomes through people, processes, and technology.
We offer the same commitment to our employees by providing professional development, meaningful projects, and flexibility to spend time with family and friends. We believe employees are at their best when fulfilled in both their professional careers and their personal lives. Learn more at voted REI Systems a Washington Post Top Workplace in 2015, 2016, 2018, 2020, 2021, 2022, 2023, 2024, and 2025!
Senior Dev Ops Engineer
REI Systems is seeking a Senior Dev Ops Engineer to lead the design, implementation, automation, and operation of enterprise-scale cloud environments and software delivery platforms supporting mission-critical federal applications.
This individual will serve as a senior technical contributor and work closely with software engineering, infrastructure, security, architecture, and operations teams to design and optimize CI/CD pipelines, cloud infrastructure, container platforms, Infrastructure as Code, automation, observability, and production operations
.
The ideal candidate brings extensive hands-on experience with AWS, Kubernetes, CI/CD, Infrastructure as Code, scripting, monitoring, production operations, and cloud-native architectures and can independently troubleshoot complex application and infrastructure issues.
Dev Ops Engineering & Automation- Design, develop, own, and continuously improve CI/CD pipelines and software delivery platforms supporting enterprise-scale applications and engineering teams.
- Architect and implement Infrastructure as Code (IaC) solutions for scalable, secure, and highly available cloud environments.
- Design, deploy, optimize, and support containerized applications and microservices using Kubernetes, Docker, AWS EKS, and/or Open Shift.
- Develop advanced automation solutions using Bash, Python, Groovy, or similar scripting languages
. - Lead application and microservice deployments across development, test, staging, and production environments.
- Partner with development teams to diagnose and resolve complex application, configuration, deployment, networking, and infrastructure issues.
- Establish and promote Dev Ops engineering standards, automation practices, and deployment best practices
. - Identify opportunities to improve cloud scalability, resiliency, security, automation, performance, and operational efficiency.
- Evaluate and implement emerging Dev Ops, cloud-native, containerization, and automation technologies
. - Provide technical guidance and mentorship to junior and mid-level Dev Ops engineers.
- Ensure the availability, reliability, scalability, and performance of 24x7 production and non-production environments
. - Lead technical response and troubleshooting efforts during complex production incidents and outages.
- Perform and facilitate root cause analysis (RCA) and develop corrective and preventative actions.
- Design and improve monitoring, alerting, logging, and observability capabilities across cloud and application environments.
- Analyze application, infrastructure, network, and system performance to proactively identify reliability and capacity issues.
- Define and track operational KPIs related to availability, uptime, SLA performance, incidents, capacity, and system health
. - Lead or oversee operational maintenance activities including operating system patching, security remediation, vulnerability management, and environment upgrades.
- Improve system observability through dashboards, metrics, alerts, distributed logging, and application performance monitoring.
- Participate in an on-call rotation and provide senior-level technical support during critical production incidents, deployments, outages, and service transitions.
- Identify recurring operational issues and develop automation and engineering solutions that reduce manual intervention.
- 7+ years of relevant Dev Ops, Cloud Engineering, Site Reliability Engineering,…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).