×
Register Here to Apply for Jobs or Post Jobs. X

Software Engineering Manager - Site Reliability Center

Job in Cleveland, Cuyahoga County, Ohio, 44101, USA
Listing for: Fairygodboss
Full Time position
Listed on 2026-09-02
Job specializations:
  • IT/Tech
    SRE/Site Reliability
Salary/Wage Range or Industry Benchmark: 100100 - 204490 USD Yearly USD 100100.00 204490.00 YEAR
Job Description & How to Apply Below

Position Overview

At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. We work together each day to foster an inclusive workplace culture where all of our employees feel respected, valued and have an opportunity to contribute to the company's success. As a(n) [position title] within PNC's [name of division] organization, you will be based in [city/state location of position].

Job

Profile

Position Overview

At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. We work together each day to foster an inclusive workplace culture where all of our employees feel respected, valued and have an opportunity to contribute to the company's success.

As a Software Engineering Manager for PNC's Site Reliability Engineering Center, you will work within PNC's Information Technology Group and be located at one of our IT Hubs:
Cleveland, Ohio;
Birmingham, Alabama;
Pittsburgh, Pennsylvania;
Dallas, Texas;
Denver, Colorado or Phoenix, Arizona and manage the daylight shift.

The Site Reliability Center (SRC) is focused on establishing a culture of operational excellence by ensuring infrastructure, platforms, and applications adhere to SRC onboarding standards that improve reliability, enable proactive issue resolution, and reduce customer impact. This role supports the vision of building a collaborative technology organization across application, infrastructure, and security teams to deliver a stable, reliable, and secure environment.

Key responsibilities include driving customer‑centric service improvements, implementing proactive and preventative reliability practices, fostering cross‑functional collaboration, enhancing monitoring and observability capabilities, promoting a blameless culture of continuous learning, and reducing operational toil through automation. The ideal candidate will help improve service performance, strengthen operational resiliency, and advance automation and observability initiatives that enhance the overall customer experience.

As a Software Engineering Manager - Site Reliability Engineering (SRE), you will lead a team responsible for ensuring the reliability, scalability, and operational excellence of mission‑critical platforms that power PNC's digital experiences. This role blends technical leadership, hands‑on problem solving, and people management, driving both production stability and continuous improvement across complex distributed systems. You will….

  • Manage SRE and related Teams; lead, coach, and develop a team of SRE engineers; set clear goals, drive accountability, and foster a culture of ownership and excellence; partner with cross‑functional stakeholders to align technology and business objectives; support talent development, performance management, and succession planning; encourage innovation, continuous learning, and Dev Ops/SRE best practices.
  • Provide after‑hours operational leadership and on‑call support. Participate in an on‑call leadership rotation supporting critical production services, major incidents, and high‑severity customer‑impacting events. Availability outside standard business hours, including evenings, weekends, and holidays, may be required to support incident response, change events, escalations, and business continuity needs.
  • Lead incident management & remediation; manage and actively participate in end‑to‑end incident response for major (P1/P2) incidents; guide real‑time triage, diagnostics, and troubleshooting across application, infrastructure, and network layers; ensure rapid execution of remediation actions and service restoration; provide clear, timely communication to stakeholders during incidents; oversee post‑incident analysis, reporting, and documentation to drive improvements.
  • Provide technical leadership in production support; serve as an escalation point for complex production issues; guide troubleshooting across: applications, infrastructure (Linux/Windows), databases (Oracle, SQL), middleware and integrations;…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary