Senior Site Reliability Engineer
Listed on 2026-07-22
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, IT Support, Systems Engineer
Job Description
At Boeing, we innovate and collaborate to make the world a better place. We’re committed to fostering an environment for every teammate that’s welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us.
The Boeing Company is looking for a Senior
Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO.
We are seeking a highly talented, motivated, and creative individual to operate, improve, and sustain mission-critical developer platforms used by Air Dominance engineering teams.
This role will provide hands-on technical ownership for Git Lab, Git Lab CI/CD runners, Jira, Confluence, PostgreSQL, and related software delivery tools such as Artifactory and Sonar Qube. The selected candidate will drive reliability improvements, automate operational workflows, troubleshoot complex incidents, lead planned maintenance activities, and help establish mature Site Reliability Engineering practices for the team.
Position Responsibilities:
Operate and maintain Git Lab, Git Lab CI/CD runners, Jira, Confluence, PostgreSQL, and related developer tooling infrastructure
Lead the deployment and configuration of software development technologies, including build servers, version control systems, CI/CD pipelines, and automated testing frameworks
Lead administration of cloud-based and on-premises infrastructure using approved Amazon Web Services (AWS), Microsoft Azure, Linux, virtualization, or container platform capabilities
Serve as a technical owner for platform reliability, availability, performance, capacity, backup, recovery, and operational readiness
Lead troubleshooting for complex application, database, runner, pipeline infrastructure, network, storage, and performance issues
Develop and maintain Infrastructure as Code (IaC), Ansible, and other automation for provisioning, configuration, platform scaling, health checks, reporting, backup validation, and routine operational tasks
Plan and execute approved changes, including application upgrades, security patches, database maintenance, runner lifecycle activities, and infrastructure updates
Establish and improve monitoring, alerting, dashboards, SLIs, SLOs, SLAs, KPIs, error budgets, and operational metrics
Lead software development tool administration, maintenance, version upgrades, patch management, and integration between tools such as Jira, Git Lab, Artifactory, Confluence, and Sonar Qube
Define, collect, analyze, and refine software delivery and platform reliability metrics to support data-driven decision making
Support incident response, root cause analysis, corrective action tracking, and post-incident reviews
Mentor junior engineers and provide technical guidance on SRE practices, secure administration, automation, and troubleshooting
Partner with developers, project administrators, cybersecurity personnel, infrastructure teams, database administrators, and program stakeholders
Improve runbooks, standard operating procedures, architecture documentation, and disaster recovery procedures
Evaluate platform risks, capacity trends, recurring incidents, and operational toil, then recommend and implement improvements
Lead demonstrations, monitor progress, and present technical status to customers and management
Lead process improvement efforts that help operationally field higher-quality end-to-end system software more frequently
Participate in after-hours support for urgent or mission-impacting issues as required
This position is expected to be 100% onsite. The selected candidate will be required to work onsite at one of the listed location options.
Basic Qualifications (Required Skills/ Experience):
Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.)
Ability to obtain access to Special Access Programs (SAP)
Bachelor's Degree
9+ years of experience with Dev Ops, Site Reliability Engineering, software engineering, and/or cloud engineering
Experience with Git Lab, Azure Dev Ops and CI/CD (Continuous Integration and Continuous Delivery (CI/CD)
Experience with technical leadership
Experience…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).