×
Register Here to Apply for Jobs or Post Jobs. X

Manager, Site Reliability Engineering

Job in Denver, Denver County, Colorado, 80285, USA
Listing for: Litera Group
Full Time, Part Time position
Listed on 2026-07-18
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 120000 - 160000 USD Yearly USD 120000.00 160000.00 YEAR
Job Description & How to Apply Below
## Manager, SREApplylocations:
Denver, COtime type:
Full time posted on:
Posted Todayjob requisition :
R-501325
** Job Description
**** Ready to Help Shape the Future of Legal Tech?!
** At Litera, we don’t just build software, we transform how the world’s top law firms operate. Every day, we Raise The BarTM️ for what’s possible through AI, innovation, and solutions that power millions of legal professionals worldwide. If you’re energized by scale, real impact, and meaningful challenges, you’ll feel right at home here.
** Where You’ll Work
** This is a hybrid role based in Denver, CO with the expectations to be in office at least 3 days a week for collaboration and connection.
** Why this Role Matters
** The Manager, Site Reliability Engineering plays a critical role in ensuring Litera’s platforms remain reliable, scalable, secure, and high performing for our customers. This role helps reduce operational risk, improve service availability, and strengthen the systems that support Litera’s continued growth. By leading a team of SREs and partnering closely across Engineering, Product, Security, and IT, this leader will drive meaningful improvements in platform resilience, incident response, automation, and operational excellence.

The impact of this role is directly tied to customer trust, system stability, and Litera’s ability to deliver dependable SaaS solutions at scale.
** What You’ll Deliver
*** Lead and develop a high-performing SRE team, creating a culture of ownership, collaboration, accountability, and continuous improvement.
* Improve the reliability, availability, and performance of Litera’s cloud and on-premises platforms through strong operational leadership and technical oversight.
* Drive effective incident response practices, including major incident leadership, escalation management, root cause analysis, and post-incident improvement.
* Establish and advance reliability metrics, including SLOs, SLIs, platform health indicators, dashboards, and reporting that improve visibility and decision-making.
* Reduce operational toil by identifying and implementing automation opportunities across runbooks, monitoring, diagnostics, infrastructure, and support processes.
* Partner with Engineering teams to improve application reliability, observability, scalability, and operational readiness across customer-facing systems.
* Support platform growth through capacity planning, performance optimization, disaster recovery readiness, and cloud/infrastructure cost optimization.
* Collaborate cross-functionally with Product, Engineering, Security, and IT to address operational risks, strengthen service delivery, and align reliability initiatives to business priorities.

We’re committed to creating an inclusive environment. If you need accommodations at any point in the process or in the role, we’re here to support you.
** What You’ll Bring
**** Must-Haves:
*** 6+ years of experience operating and troubleshooting distributed SaaS applications across hybrid environments, including on-premises infrastructure and Azure (preferred) or AWS cloud platforms.
* 2+ years of SRE management experience leading teams that support multi-product platforms, including experience with resource planning, workload prioritization, team development, and scaling teams.
* Strong leadership experience managing major incidents, serving as an escalation point, leading blameless RCAs, and driving continuous post-incident improvements.
* Deep experience with observability, monitoring, alerting, APM tools, and reliability metrics, including SLOs, SLIs, and platform health reporting.
* Strong technical background in infrastructure automation, configuration management, CI/CD practices, and tools such as Terraform, Ansible, or similar technologies.
* Excellent communication, prioritization, and cross-functional collaboration skills, with the ability to align stakeholders and guide teams through complex operational challenges.
** Nice to Haves:
*** Previous experience in a Site Reliability Engineering function within a SaaS organization.
* Familiarity with ITIL, Problem Management, Change Management, Operational Excellence, or modern platform engineering…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary