Senior Platform Site Reliability Engineer
Listed on 2026-09-06
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Job Details
Primary
Skills:
SRE principles (advanced), Observability (advanced), Incident Response (advanced), SaaS operations (advanced), AI tooling (intermediate)
Contract Type: W2 Only
Duration: 12+ Months
Location:
Lehi, UT (#LI - Hybrid)
Pay Range: $50-$55/Hr. on W2
Job Summary:
This role focuses on maintaining the stability and availability of SaaS solutions, including both internally managed platforms and third-party services. You will leverage SRE best practices to ensure operational health, even when full control over the technology stack is not possible. Success will involve proactive monitoring, effective incident management, and strong collaboration with various teams and vendors to drive reliability improvements.
- Ensure the uptime and performance of SaaS solutions.
- Lead incident response and post-incident analysis.
- Develop robust observability for partially owned systems.
- Drive improvements based on operational data and learnings.
- Collaborate with teams and vendors for reliable operations.
- Proven experience in Site Reliability Engineering.
- Hands-on experience operating SaaS or third-party systems.
- Strong incident management and communication skills.
Industry
Experience:
Experience in managing and ensuring the reliability of Software as a Service (SaaS) offerings, particularly those involving vendor-managed components, is highly preferred.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).