Senior Site Reliability Engineer
Listed on 2026-08-26
-
IT/Tech
Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Systems Administrator
At Boeing, we innovate and collaborate to make the world a better place. We’re committed to fostering an environment for every teammate that’s welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us.
At Boeing, we innovate and collaborate to make the world a better place. We’re committed to fostering an environment for every teammate that’s welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us.
The Boeing Company is looking for a Senior Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO.
We are seeking a highly talented, motivated, and creative individual to operate, improve, and sustain mission‑critical developer platforms used by Air Dominance engineering teams.
This role will provide hands‑on technical ownership for Git Lab, Git Lab CI/CD runners, Jira, Confluence, PostgreSQL, and related software delivery tools such as Artifactory and Sonar Qube. The selected candidate will drive reliability improvements, automate operational workflows, troubleshoot complex incidents, lead planned maintenance activities, and help establish mature Site Reliability Engineering practices for the team.
Position Responsibilities- Operate and maintain Git Lab, Git Lab CI/CD runners, Jira, Confluence, PostgreSQL, and related developer tooling infrastructure
- Lead the deployment and configuration of software development technologies, including build servers, version control systems, CI/CD pipelines, and automated testing frameworks
- Lead administration of cloud‑based and on‑premises infrastructure using approved Amazon Web Services (AWS), Microsoft Azure, Linux, virtualization, or container platform capabilities
- Serve as a technical owner for platform reliability, availability, performance, capacity, backup, recovery, and operational readiness
- Lead troubleshooting for complex application, database, runner, pipeline infrastructure, network, storage, and performance issues
- Develop and maintain Infrastructure as Code (IaC), Ansible, and other automation for provisioning, configuration, platform scaling, health checks, reporting, backup validation, and routine operational tasks
- Plan and execute approved changes, including application upgrades, security patches, database maintenance, runner lifecycle activities, and infrastructure updates
- Establish and improve monitoring, alerting, dashboards, SLIs, SLOs, SLAs, KPIs, error budgets, and operational metrics
- Lead software development tool administration, maintenance, version upgrades, patch management, and integration between tools such as Jira, Git Lab, Artifactory, Confluence, and Sonar Qube
- Define, collect, analyze, and refine software delivery and platform reliability metrics to support data‑driven decision making
- Support incident response, root cause analysis, corrective action tracking, and post‑incident reviews
- Mentor junior engineers and provide technical guidance on SRE practices, secure administration, automation, and troubleshooting
- Partner with developers, project administrators, cybersecurity personnel, infrastructure teams, database administrators, and program stakeholders
- Improve runbooks, standard operating procedures, architecture documentation, and disaster recovery procedures
- Evaluate platform risks, capacity trends, recurring incidents, and operational toil, then recommend and implement improvements
- Lead demonstrations, monitor progress, and present technical status to customers and management
- Lead process improvement efforts that help operationally field higher‑quality end‑to‑end system software more frequently
- Participate in after‑hours support for urgent or mission‑impacting issues as required
This position is expected to be 100% onsite. The selected candidate will be required to work onsite at one of the listed location options.
Basic Qualifications (Required Skills/ Experience)- Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.)
- Ability to obtain access to Special Access Programs (SAP)
- Bachelor's Degree
- 9+ years of experience…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).