×
Register Here to Apply for Jobs or Post Jobs. X

Lead Site Reliability Engineer

Job in Berkeley, St. Louis County, Missouri, USA
Listing for: Boeing
Full Time position
Listed on 2026-07-22
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Cybersecurity
Job Description & How to Apply Below

Job Description

At Boeing, we innovate and collaborate to make the world a better place. We’re committed to fostering an environment for every teammate that’s welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us.

The Boeing Company is looking for a Lead
Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO
.

We are seeking a highly talented, motivated, and creative technical leader responsible for the reliability strategy, architecture, operational maturity, and long-term technical direction of mission-critical developer platforms used by Air Dominance engineering teams.

This role will provide technical leadership across Git Lab, Git Lab CI/CD runners, Jira, Confluence, PostgreSQL, related software delivery tools such as Artifactory and Sonar Qube, and the supporting infrastructure, automation, monitoring, backup, recovery, and security controls required to operate these services. The selected candidate will define standards, guide architecture decisions, mentor engineers, lead complex technical investigations, and partner with program leadership and stakeholders to ensure developer tooling remains secure, reliable, scalable, and supportable.

Position Responsibilities:

  • Define and lead the Site Reliability Engineering technical strategy for Git Lab, CI/CD runners, Jira, Confluence, PostgreSQL, Artifactory, Sonar Qube, and related developer tooling infrastructure

  • Establish platform reliability architecture, operational standards, SLIs, SLOs, SLAs, KPIs, error budgets, observability patterns, capacity models, backup strategies, and disaster recovery approaches

  • Serve as the senior technical authority for complex reliability, performance, scalability, integration, database, automation, and security-related platform decisions

  • Lead architecture and design reviews for developer tooling infrastructure, CI/CD runner topology, PostgreSQL operations, cloud-based and on-premises infrastructure, monitoring, alerting, access controls, and platform integrations

  • Drive automation, Infrastructure as Code, Ansible, configuration management, and repeatable operational patterns that reduce toil and improve reliability

  • Guide major upgrades, migrations, lifecycle planning, patch strategies, recovery planning, and technical roadmaps for supported platforms

  • Lead the most complex incidents and technical investigations, including root cause analysis, corrective action planning, and systemic reliability improvements

  • Mentor and technically guide SREs in operational excellence, troubleshooting, automation, secure administration, and architectural thinking

  • Partner with program leadership, cybersecurity, infrastructure, software engineering, database, networking, suppliers, customers, and other stakeholders

  • Identify platform risks, technical debt, capacity constraints, single points of failure, compliance concerns, and operational gaps, then drive remediation plans

  • Define, collect, analyze, and refine software delivery and platform reliability metrics for team execution, management visibility, and continuous improvement

  • Lead standards for provisioning, platform scaling, configuration management, monitoring, troubleshooting, and software delivery tool integration

  • Develop and maintain architecture documentation, design patterns, standards, operational readiness criteria, and executive technical briefings

  • Influence support models, maintenance strategies, escalation paths, and investment priorities based on mission impact and operational risk

  • Lead efforts to operationally field higher-quality end-to-end system software more frequently

  • Participate in after-hours support and escalation for urgent or mission-impacting issues as required

This position is expected to be 100% onsite. The selected candidate will be required to work onsite at one of the listed location options.

Basic Qualifications (Required Skills/ Experience):

  • Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.)

  • Ability to obtain access to Special Access Programs (SAP)

  • Bachelor's Degree

  • 14+ years of…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary