Reliability/Systems Engineer - On W2 - Onsite at MN/WS/TX/OR states
Listed on 2026-08-05
-
IT/Tech
SRE/Site Reliability, IT Support, Systems Engineer, Cloud Computing: Infrastructure & Operations
Reliability/Systems Engineer
Position:
Reliability/Systems Engineer
Location:
Saint Paul, MN (No remote – must be onsite at one of the following: Minnesota, Dallas TX, Brookfield WI, Portland OR) On W2 contract only.
Required Skills:
Strong observability experience across APM, logs, monitoring, alerts, and dashboards. Proficiency with tools:
Datadog, App Dynamics, Splunk, Service Now. Ability to support non-prod environments with strict SLAs and incident response.
Job Description:
Very strong candidate with observability experience. Experience should include full product experience (APM, Logs, setting up monitoring, alerts, dashboards). Overall looking for a good Reliability Engineer that will support our non-prod environments by setting up alerting, monitoring strict SLA's and engaging to determine issues. Must have knowledge of the following tools:
Datadog, App Dynamics, Splunk Logging, and Service Now for reviewing incidents, problem records, and reporting. Strong analytics skills for researching incident rates and noise.
Project
Summary:
This project focuses on establishing monitoring and availability for non-production environments with tight SLAs.
Top Responsibilities:
- Monitor systems using observability tools
- Respond to issues and engage teams for resolution
- Report on SLA adherence
- Participate in on-call rotation
- Analyze incident trends and reduce noise
Required Skills (3+ years each):
- Splunk
- App Dynamics
- Datadog
- Batch Monitoring
- Service Now
Hard Requirements:
- Datadog
- Splunk
- Understanding DB, server, app monitoring process and how remediation efforts begin (not doing the remediation)
D2D:
Reliability engineer
- First shift w/ some on call – divided between 3 individuals during day shift
- Observability, app monitoring, ticket intake, etc for app support
- Must have knowledge of the following tools:
Datadog, App Dynamics, Splunk Logging, and Service Now - Creating visibility to resources and teams outside of the production environments (“Non-Prod”)
- Working alongside of dev teams for API and App build to document their process
- Creating dashboards and using Jira for backlog but main source of daily work comes from collab with Dev Teams
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).