More jobs:
Monitoring Engineer
Job in
Sioux Falls, Minnehaha County, South Dakota, 57102, USA
Listed on 2026-07-25
Listing for:
TEKsystems c/o Allegis Group
Full Time
position Listed on 2026-07-25
Job specializations:
-
IT/Tech
SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations, IT Support
Job Description & How to Apply Below
Our client, a leader in the healthcare industry, is seeking a Senior Monitoring Engineer to support and enhance enterprise monitoring, observability, and alerting solutions across a large and complex technology environment. This role will partner closely with infrastructure, networking, application, and engineering teams to design and implement monitoring strategies that improve system reliability, reduce downtime, and strengthen incident response capabilities.
The ideal candidate will bring deep experience with enterprise monitoring platforms, a proactive approach to problem-solving, and a passion for building scalable observability solutions that support mission-critical operations.
Project Overview
As a Senior Monitoring Engineer, you will play a key role in improving visibility across critical healthcare systems, applications, infrastructure, databases, and networks. This position will focus on enhancing enterprise monitoring capabilities, implementing modern observability tools, and developing proactive solutions that help identify issues before they impact patient care or business operations.
You will help drive monitoring standardization efforts, support platform modernization initiatives, and collaborate with teams across the organization to ensure monitoring and alerting are built into all new technology deployments.
Key Responsibilities
- Design, deploy, and support enterprise monitoring and alerting solutions across infrastructure, applications, databases, and networks
- Develop monitoring strategies that improve system availability, reliability, and performance
- Implement and maintain dashboards that provide operational visibility across multiple platforms and teams
- Ensure accurate collection, retention, and reporting of metrics for alerting, analysis, and trend reporting
- Collaborate with engineering teams to integrate monitoring requirements into new deployments and technology initiatives
- Build and maintain reporting solutions that provide insights into operational health and incident trends
- Support incident management processes by ensuring monitoring and alerting systems align with response protocols
- Identify recurring issues and develop proactive solutions to improve system stability
- Maintain platform health through upgrades, patching, troubleshooting, and vendor engagement when necessary
- Assist in developing automation, scripting, and self-healing capabilities within the environment
- Mentor and provide technical guidance to fellow engineers and team members
- Improve enterprise-wide observability and monitoring maturity
- Increase proactive detection of issues before they impact patient care
- Strengthen incident response and operational awareness across IT teams
- Enhance visibility into mission-critical healthcare systems and infrastructure
- Drive standardization of monitoring tools, processes, and reporting
- Support the reliability and availability of systems serving approximately 2.4 million patients
- Contribute to the continued growth and modernization of a large healthcare technology environment
- Experience designing, implementing, and supporting enterprise monitoring and observability solutions
- Strong knowledge of network, infrastructure, application, and systems monitoring
- Experience deploying new monitoring tools and improving enterprise-wide observability practices
- Experience with log aggregation or SIEM platforms such as Splunk, Log Rhythm, Graylog, ELK, or similar solutions
- Strong troubleshooting and analytical skills
- Experience working with incident management and operational response processes
- Ability to communicate effectively with technical teams and leadership stakeholders
- Experience creating dashboards, reporting solutions, and performance metrics
- Strong documentation and process improvement experience
- Experience with Logic Monitor, Thousand Eyes, Big Panda, Service Now AIOps, or similar observability platforms
- Experience supporting tools such as Dynatrace, Pager Duty, Victor Ops, Solar Winds, Nagios, Zabbix, Datadog, App Dynamics, or New Relic
- Scripting and automation experience
- Experience developing…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×