More jobs:
Major Incident Management & NOC Lead
Job in
Wilmington, New Castle County, Delaware, 19894, USA
Listed on 2026-08-29
Listing for:
Bi3 Careers
Full Time
position Listed on 2026-08-29
Job specializations:
-
IT/Tech
Systems Administrator, SRE/Site Reliability, IT Project Manager, Cloud Computing: Infrastructure & Operations
Job Description & How to Apply Below
We are seeking an experienced Major Incident Management & NOC Lead to lead enterprise NOC operations and manage critical P1/P2 incidents. This person will serve as the Incident Commander during major outages, coordinate technical teams and vendors, communicate with leadership, and drive incidents through resolution and root cause analysis.
The ideal candidate has a strong infrastructure/network operations background along with proven NOC and major incident leadership.
Key Responsibilities- Lead NOC operations, monitoring, escalation and operational processes
- Serve as Incident Commander for P1/P2 incidents and major outages
- Coordinate network, infrastructure, application, security, cloud and vendor teams
- Provide clear incident updates to business and executive leadership
- Lead post-incident reviews, root cause analysis and corrective actions
- Manage vendor escalations and SLA performance
- Maintain SOPs, runbooks and escalation procedures
- Drive improvements in monitoring, alerting and service reliability
- 10+ years in IT Operations, NOC and/or Major Incident Management
- Previous experience leading a NOC or enterprise operations team
- Strong P1/P2 incident management and outage leadership experience, including post-incident reviews, root cause analysis and problem management
- Strong knowledge of ITIL Incident, Problem and Change Management
- Experience with Service Now or a similar ITSM platform
- Experience with SLAs, SOPs, escalation procedures and vendor management
- Strong hands‑on technical foundation in infrastructure and network operations, with the ability to lead network, Linux/Windows and operations teams
- Strong infrastructure knowledge including Windows/Linux, DNS, DHCP, TCP/IP, routing, load balancers and firewalls
- Experience with AWS and/or Azure
- Experience with monitoring tools such as Splunk, Dynatrace, Datadog, New Relic, App Dynamics or similar
- Strong leadership, problem-solving and communication skills
- Ability to remain calm and make decisions during critical outages
- Bachelor’s degree or equivalent experience
- ITIL certification
- Power Shell, Python or Bash experience
- Experience in a large, high-availability enterprise environment
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×