×
Register Here to Apply for Jobs or Post Jobs. X

Technical Enterprise Incident Manager

Remote / Online - Candidates ideally in
Linthicum, Anne Arundel County, Maryland, USA
Listing for: Peraton
Remote/Work from Home position
Listed on 2026-10-03
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, IT Support
Salary/Wage Range or Industry Benchmark: 120000 - 160000 USD Yearly USD 120000.00 160000.00 YEAR
Job Description & How to Apply Below
Basic Qualifications:
  • Must be a U.S. citizen with the ability to obtain and maintain the required Public Trust level clearance
  • Bachelor’s Degree and 5 years of experience, or a High School diploma or equivalent and 9 years of experience
  • 5+ years of experience in Cloud Incident Management, Operations Engineering, NOC, SRE, Application or Production Support environments.
  • Experience leading enterprise Major Incident response efforts in a 24x7 operational environment.
  • Strong understanding of ITIL Incident and Problem Management processes.
  • 3+ years of experience working with AWS cloud services
  • Experience with monitoring and observability platforms such as Datadog, Cloud craft, or similar
  • Experience using Service Now or similar ITSM platforms.
  • Strong analytical, troubleshooting, and organizational skills.
  • Excellent written and verbal communication skills with ability to facility meetings as well as brief technical teams and executive leadership.
Preferred Qualifications:
  • Experience in a Site Reliability Engineering (SRE) or Cloud Platform Dev Ops environment.
  • Experience supporting federal, healthcare, financial, or other highly regulated environments.
  • Hands-on experience with infrastructure technologies including:
    • Windows/Linux Servers
    • Networking concepts
    • Cloud platforms (AWS, Azure, or GCP)
    • Load balancers, proxies, DNS, and firewalls

Peraton is seeking a highly motivated and technically skilled Technical Enterprise Incident Manager with strong Cloud Platform and application experience to lead enterprise incident response, service restoration efforts, and operational reliability initiatives. This individual will serve as the central point of coordination during major incidents, ensuring rapid resolution, clear communication, and continuous service improvement across enterprise infrastructure and applications.

The ideal candidate possesses a strong operational background, excellent communication skills, and hands‑on technical expertise in infrastructure, cloud technologies, monitoring, automation, and IT service management processes. This role requires the ability to drive incident response while also identifying systemic reliability improvements.

Location:

Remote Shift

Schedule:

8am – 5pm Eastern Standard Time (EST). This position will also participate in 24x7 on‑call rotations for incident management.

What You Will Do Enterprise Incident Management
  • Lead and coordinate Incident bridge calls involving infrastructure, application, network, cloud, security, and vendor teams.
  • Drive rapid service restoration while maintaining accurate timelines, communications, and executive updates.
  • Ensure incidents are prioritized appropriately based on business impact and operational risk.
  • Manage escalation procedures and engage leadership when required.
  • Monitor SLA compliance and ensure incident response metrics are consistently achieved.
  • Facilitate Post‑Incident Reviews (PIRs) and ensure high‑quality Root Cause Analysis (RCA) documentation and follow‑through on corrective actions.
  • Drive automation of incident detection, triage, and response workflows to reduce operational toil and improve resolution times.
  • Maintain structured communication frameworks during major incidents, including timely stakeholder notifications, ongoing updates, and final incident summaries.
  • Validate runbooks, service dependency maps, and technical documentation to ensure accuracy and usability during incidents.
  • Utilize additional observability tools such as log aggregation and APM to enhance troubleshooting and incident analysis.
Cloud Platform Dev Sec Ops  Engineering
  • Work with application teams to facilitate issues and implement root cause remediations.
  • Develop and enhance monitoring, alerting, and dashboarding capabilities.
  • Analyze trends, KPIs, and operational…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary