×
Register Here to Apply for Jobs or Post Jobs. X

Manager, Site Reliability Engineering

Job in Buffalo, Erie County, New York, 14201, USA
Listing for: M&T Bank
Full Time position
Listed on 2026-09-05
Job specializations:
  • IT/Tech
    SRE/Site Reliability, IT Project Manager, Systems Engineer, Cloud Computing: Infrastructure & Operations
Job Description & How to Apply Below

Manager, Site Reliability Engineering

Responsible for leading the Site Reliability Engineering Center of Excellence and the Forward Deployed SRE program supporting critical banking platforms, applications, and technology services. Manages an organization of employees and contingent resources through direct reports, program managers, SRE leaders, technical leads, and matrixed delivery relationships.

Establishes the enterprise SRE strategy, operating model, engineering standards, governance, talent model, and adoption roadmap. Accountable for improving service reliability, availability, scalability, performance, resiliency, deployment safety, and operational maturity across supported technology domains.

Leads the deployment of SRE capabilities into application and platform teams through a Forward Deployed SRE model. Partners with senior leaders across application development, infrastructure, cloud engineering, architecture, cybersecurity, technology operations, risk, and business-aligned technology organizations to prioritize engagements and deliver measurable reliability improvements.

Balances strategic leadership, people management, program execution, technical governance, and operational accountability. Ensures SRE practices are implemented consistently and that reliability investments produce measurable improvements in customer experience, operational risk, engineering productivity, release quality, and service performance.

Primary Responsibilities

SRE Strategy and Center of Excellence Leadership

  • Establish and execute the vision, strategy, operating model, service offerings, and multiyear roadmap for the Site Reliability Engineering Center of Excellence.
  • Define enterprise SRE standards, engineering practices, governance processes, engagement models, and maturity expectations.
  • Lead the adoption of reliability engineering practices across application development, infrastructure, platform engineering, cloud engineering, and technology operations.
  • Translate enterprise technology, business, resiliency, and risk priorities into an actionable SRE portfolio and delivery roadmap.
  • Establish a scalable SRE service model that includes consulting, enablement, embedded engineering, Forward Deployed SRE engagements, reusable capabilities, and sustained ownership by application and platform teams.
  • Define intake, prioritization, engagement, transition, and exit criteria for SRE services.
  • Develop and maintain an SRE maturity model used to assess service teams, identify reliability gaps, and guide improvement plans.
  • Establish communities of practice, technical forums, training programs, playbooks, reference architectures, and reusable engineering patterns that expand SRE capabilities across the organization.
  • Ensure the SRE Center of Excellence remains aligned with enterprise engineering standards, cloud strategy, operational risk requirements, and evolving industry practices.
  • Represent the SRE organization in senior leadership forums, architecture reviews, operational governance meetings, and enterprise transformation initiatives.

Forward Deployed SRE Program Leadership

  • Lead the Forward Deployed SRE program, placing SRE professionals into high-priority application and platform teams to address complex reliability challenges and improve operational maturity.
  • Manage program managers, SRE leaders, and technical leads responsible for coordinating engagements across multiple technology domains.
  • Establish a transparent intake and prioritization process based on customer impact, service criticality, operational risk, incident history, reliability maturity, strategic importance, and anticipated business value.
  • Partner with application and platform leaders to define engagement objectives, scope, deliverables, staffing, success measures, dependencies, and duration.
  • Ensure Forward Deployed SRE teams deliver sustainable engineering improvements rather than becoming long-term substitutes for application support or production operations.
  • Establish shared accountability for participation, knowledge transfer, remediation activities, and long-term ownership of implemented reliability capabilities.
  • Develop transition and exit plans that enable application and platform teams to sustain SRE practices after engagements conclude.
  • Evaluate engagement effectiveness using measurable outcomes such as availability, SLO attainment, incident frequency, restoration time, alert quality, automation adoption, toil reduction, change-failure rate, deployment reliability, and engineering maturity.
  • Convert common findings and lessons learned into reusable standards, automation, tools, training, and engineering patterns.
  • Continuously optimize the Forward Deployed SRE operating model based on demand, capacity, outcomes, stakeholder feedback, and changes in technology strategy.

Organizational and People Leadership

  • Lead an organization of employees and contingent resources through direct and indirect management relationships.
  • Manage and develop program managers, SRE managers, technical…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary