×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer; SRE Actimize Fraud Platforms

Remote / Online - Candidates ideally in
Buffalo, Erie County, New York, 14202, USA
Listing for: M&T Bank
Remote/Work from Home position
Listed on 2026-10-08
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Job Description & How to Apply Below
Position: Site Reliability Engineer (SRE) – NICE Actimize Fraud Platforms
This role is four days onsite at our Seneca One Buffalo, NY location, with the flexibility to work from home one day per week Position Summary Responsible for the reliability, availability, performance, and operational excellence of critical NICE Actimize Fraud and Financial Crime platforms. Leads production support, release engineering, observability, automation, and cloud transformation initiatives. Partners with application, infrastructure, and business teams to improve platform resiliency, streamline software delivery, eliminate manual operational tasks, and support strategic initiatives including AIQ adoption and data center/Cloud migration programs.

Key Responsibilities:

Provide production support for NICE Actimize Fraud, AML, and Financial Crime applications, ensuring high availability and system stability.

Lead incident management, root cause analysis (RCA), problem management, and service restoration activities.

Develop and maintain observability solutions using Dynatrace, Open Telemetry (OTel), logging, tracing, dashboards, and automated alerting.

Build and support CI/CD pipelines to enable reliable, repeatable, and automated application deployments.

Drive automation initiatives to eliminate manual, repetitive operational tasks and improve operational efficiency.

Implement Infrastructure as Code (IaC) using Terraform and automation tooling to standardize environment provisioning and deployment processes.

Support release engineering activities, deployment governance, release validation, and post-deployment monitoring.

Partner with development teams to improve application reliability, performance, scalability, and operational readiness.

Design and implement automated regression, smoke, and release validation testing.

Support Microsoft Azure environments and cloud-native operational tooling.

Lead and support AIQ platform onboarding, integration, monitoring, and operational support initiatives.

Participate in and support Colo migration and Azure migration activities, including infrastructure readiness, deployment automation, application validation, production cutovers, and post-migration stability.

Define and monitor SLIs, SLOs, and operational metrics to improve service reliability and customer experience.

Create and maintain operational runbooks, support procedures, and technical documentation.

Education and Experience

Required:

Combined minimum of 8 years’ higher education and/or work experience in systems design, management and/or architecture NICE Actimize product family support experience and integrations. Required skills include mainframe experience and proficiency in running DBA/SQL queries to validate transaction health and integrity.

Strong experience with Dynatrace, Open Telemetry, monitoring, logging, alerting, and observability platforms.

Hands-on experience with CI/CD pipelines, release engineering, and deployment automation.

Terraform and Infrastructure as Code (IaC) expertise.

Experience automating operational workflows and eliminating manual support processes.

Strong knowledge of incident management, problem management, and production support practices.

Experience with Microsoft Azure and cloud operations.

Familiarity with AIQ platform support and large-scale infrastructure migration initiatives.

Understanding of SRE principles, reliability engineering, and operational excellence.

Required skills include mainframe experience and proficiency in running DBA/SQL queries to validate transaction health and integrity.

Proven expertise in Production Support, Site Reliability Engineering (SRE), Incident Management, Release Engineering, and Operational Excellence.

Strong experience managing large-scale production environments, ensuring high availability, resiliency, performance, and service…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary