×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Champion

Job in Wayne, Delaware County, Pennsylvania, 19087, USA
Listing for: SwiftCruit
Full Time position
Listed on 2026-07-18
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 120000 - 190000 USD Yearly USD 120000.00 190000.00 YEAR
Job Description & How to Apply Below

Core Responsibilities

  • Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.

  • Design and implement processes that enforce enterprise resiliency and reliability standards.

  • Lead blameless post‑incident reviews for high‑severity incidents or incidents spanning multiple complex product families.

  • Partner with product and platform teams to proactively identify and remediate reliability risks before they impact clients.

  • Develop, communicate, and evangelize new standards, tools, and frameworks across subdivisions, ensuring consistent adoption.

  • Troubleshoot complex production issues and implement durable solutions that prevent recurrence.

  • Participate in a periodic on‑call rotation to support production stability.

  • Evaluate and onboard resiliency and reliability tooling.

  • Actively participate in reliability engineering and resilience communities of practice, contributing to shared learning and enterprise consistency.

  • Contribute to strategic initiatives that advance Vanguard’s operational maturity and resiliency posture.

Qualifications | Technical Skills
  • Observability Platforms:
    Experience with modern observability and monitoring tools, such as Splunk, Honeycomb, Cloud Watch, Dynatrace, or App Dynamics.

  • Reliability Metrics:
    Strong understanding of SLIs, SLOs, and SLAs, including dashboarding and reporting practices.

  • Monitoring & Alerting:
    Experience with alert design, anomaly detection, predictive alerting, and synthetic monitoring using structured methodologies.

  • Automation & Resilience Engineering:
    Experience with automation and resilience practices such as Python-based automation, RPA platforms (e.g., Blue Prism, UiPath), chaos engineering, and failure analysis techniques (e.g., FMEA).

Special Factors

Vanguard is not offering visa sponsorship for this position.

#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary