×
Register Here to Apply for Jobs or Post Jobs. X

Technology Engineering - Lead Platform Engineer; SRE); IC

Job in Evansville, Natrona County, Wyoming, 82636, USA
Listing for: Mindlance
Full Time position
Listed on 2026-09-25
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Systems Engineer, Cloud Computing: Infrastructure & Operations, Cybersecurity
Job Description & How to Apply Below
Position: Technology Engineering - Lead Platform Engineer (SRE) (IC3)
Location: Evansville

Job-ID Reference
26-26649

Remote
50% Remote The Lead Monitoring and Observability Engineer (IC3) serves as a senior technical contributor responsible for ensuring monitoring reliability, telemetry quality, automation maturity, and operability across OMF’s infrastructure and application ecosystem. This role acts as a technical mentor, monitoring and networking SME, and hands-on engineer who guides monitoring-platform evolution, improves service quality, and collaborates with product and engineering teams to deliver scalable, stable, and observable systems.

Key Responsibilities Monitoring Reliability & Performance
• Establish platform SLOs, availability goals, latency/error budgets, reliability metrics, and monitoring coverage expectations in partnership with teams.
• Continuously measure service and platform health and implement changes to improve reliability, performance, alert quality, and operational stability.
• Define and maintain standards for actionable alerting, dashboards, logs, metrics, traces, and service-health reporting. Architecture & Technical Leadership
• Lead observability design efforts for Elastic/ELK, telemetry pipelines, monitoring platforms, dashboards, alerting, distributed tracing, synthetic monitoring, microservices platforms, cloud infrastructure, or network monitoring, depending on assignment.
• Provide “out-of-the-box” technical solutions that balance velocity, reliability, operational visibility, and cost.
• Evaluate monitoring and observability tools, patterns, integrations, and emerging requirements.
• Define reusable observability patterns, reference architectures, onboarding guidance, and validation practices for product and platform teams. Hands-On Engineering & Automation
• Perform advanced configuration, IaC development (Terraform/Ansible), monitoring-as-code development, CI/CD pipeline engineering, and cloud platform automation.
• Build and maintain scalable, resilient monitoring and observability components that support product teams.
• Implement and maintain monitoring, alerting, data visualization, logging, metrics, tracing, and telemetry collection capabilities as noted in the IC3 role reference.
• Improve automation for monitoring onboarding, dashboard creation, alert configuration, telemetry collection, and operational workflows. Operational Excellence
• Reduce manual operations through automation and self-service observability tooling.
• Review, optimize, and maintain observability capabilities including logs, metrics, traces, dashboards, alerting, and service-health reporting.
• Improve alert signal quality, reduce alert noise, and ensure alerts have clear ownership, escalation paths, and actionable runbook guidance.
• Participate in and lead high-severity incident response for monitoring and observability-owned domains.
• Support root-cause analysis by correlating telemetry, infrastructure conditions, application behavior, and operational events. Network Monitoring
• Define and maintain network-monitoring standards for network availability, reachability, latency, packet loss, interface health, capacity, device health, routing, and dependency-related service impact.
• Partner with Networking, Cloud, App Dev, Security, and SRE teams to establish monitoring coverage for critical network devices, services, and customer journeys.
• Build and maintain network-monitoring dashboards, alerting patterns, service-health views, and operational workflows.
• Establish validation practices for network-device onboarding, telemetry collection, alert quality, dashboard completeness, and operational readiness.
• Support the migration of network monitoring from Ops Ramp to the selected network-monitoring tool, including requirements definition, technical…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary