Senior NOC AI-Ops Engineer
Listed on 2026-07-25
-
IT/Tech
SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations
Role Overview
We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational resilience, and intelligent automation initiatives across enterprise technology environments. This role will partner with Network Operations Center (NOC), Infrastructure Operations, Cloud Engineering, Dev Ops, and Application Support teams to proactively detect, respond to, and prevent technology incidents.
What You Will DoMonitor, document, and analyze major incident response efforts and service recovery activities. Serve as a senior escalation point for Tier 1 and Tier 2 operational incidents. Conduct incident reviews, root cause analysis, and corrective action planning. Improve Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
Why It Might Be a FitThe ideal candidate combines hands-on incident management expertise with experience implementing observability, automation, and AI-driven operational solutions to improve system reliability, reduce operational overhead, and enhance customer experience.
RequirementsDeep expertise in AIOps, ITSM, ITIL, SRE, Incident Management, Cloud Operations, and Enterprise Infrastructure Hands-on incident management expertise Experience implementing observability, automation, and AI-driven operational solutions
BenefitsBenefits Competitive salary
Equity
HealthcarePTORetirement
Learning budget
Parental leave
Wellness
Visa/relocation
Remote flexibility
Stipends
Bonus/commission
Paid holidays
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).