×
Register Here to Apply for Jobs or Post Jobs. X

Systems Engineer, Senior - Observability

Job in Somerville, Middlesex County, Massachusetts, 02143, USA
Listing for: Brigham and Women's Hospital
Full Time position
Listed on 2026-07-26
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations, IT Support, Cybersecurity
Job Description & How to Apply Below
Site:
Mass General Brigham Incorporated

Mass General Brigham relies on a wide range of professionals, including doctors, nurses, business people, tech experts, researchers, and systems analysts to advance our mission. As a not-for-profit, we support patient care, research, teaching, and community service, striving to provide exceptional care. We believe that high-performing teams drive groundbreaking medical discoveries and invite all applicants to join us and experience what it means to be part of Mass General Brigham.

Job Summary

Senior Observability Systems Engineer

Join Mass General Brigham's Digital Enterprise Observability team as a Senior Observability Systems Engineer. In this role, you'll help strengthen the reliability, performance, and visibility of critical digital services across the enterprise. You'll work hands-on with observability platforms such as Dynatrace and Cisco Thousand Eyes, build automation and configuration-as-code solutions, and partner with application, cloud, network, and security teams to improve how teams monitor, troubleshoot, and respond to issues.

This is a technical engineering role that combines platform engineering, automation, Dev Ops practices, and observability expertise. The position does not involve direct patient care, but the work supports the digital systems that help Mass General Brigham deliver care across the organization.

What You'll Do

* Deploy, configure, and optimize enterprise observability platforms, primarily Dynatrace, along with Cisco Thousand Eyes for network-path and digital experience monitoring.

* Build and maintain dashboards, alerts, management zones, tagging rules, synthetic tests, and network-path tests that help teams detect issues earlier and troubleshoot faster.

* Develop and maintain configuration-as-code using Terraform, with repositories and pipelines in Azure Dev Ops today and a planned move to Git Hub Enterprise.

* Create and maintain custom Dynatrace extensions to expand monitoring coverage for systems that do not have native support.

* Implement observability for Dev Ops frameworks in Azure, including instrumentation and quality gates within CI/CD pipelines.

* Onboard applications and infrastructure to observability platforms, including instrumentation, monitoring configuration, and data validation.

* Automate alerting and remediation workflows to reduce mean time to resolution and improve service uptime.

* Partner with application, cloud, network, and security teams to establish and apply observability standards across the environment.

* Create user documentation, operational guidance, and best practices to help teams use observability tools effectively.

* Use standard work, project management tools, and change management processes to deliver reliable and well-coordinated work.

* Model Mass General Brigham's values through collaboration, accountability, service commitment, innovation, integrity, respect, continuous improvement, and teamwork.

Qualifications

What You'll Bring

Education:

Bachelor's degree in Computer Science or a related field required. Relevant experience may be considered in lieu of a degree.

Experience:

5-7 years of experience as a systems engineer or in a related technical engineering role.

Required Skills and Experience

* Hands-on experience with Dynatrace, including application performance monitoring, infrastructure monitoring, real user monitoring, and Dynatrace Query Language.

* Experience with Cisco Thousand Eyes for synthetic testing, path visualization, and internet or wide area network performance monitoring.

* Experience using Terraform for configuration as code, including the Dynatrace Terraform provider.

* Experience with source control and pipelines in Azure Dev Ops; familiarity with Git Hub Enterprise is helpful as the organization transitions toward that standard.

* Strong programming skills in Python and JavaScript for automation, custom telemetry, and tooling.

* Experience developing custom Dynatrace extensions.

* Experience implementing observability for Dev Ops frameworks and CI/CD pipelines in Azure.

* Foundational knowledge of technology infrastructure, including applications, servers, storage, and networks.

* Strong analytical and troubleshooting skills across metrics, logs, and traces.

* Ability to communicate clearly with both technical and non-technical stakeholders.

* Strong collaboration skills, including experience working across teams and with vendors to achieve results.

* Ability to mentor team members and support better technical outcomes.

Preferred Qualifications

* Power Shell scripting experience.

* Familiarity with Splunk, Microsoft SCOM, Site Scope, or Nagios.

* Dynatrace Associate or Professional certification.

* Azure certification, such as Azure Administrator or Azure Dev Ops Engineer.

* Experience with Monaco and Dynatrace Grail.

* Exposure to Site Reliability Engineering practices, including service level indicators, service level objectives, and error budgets.

Additional Job Details (if applicable)

* Full-time…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary