Head of Site Reliability Engineering; SRE
Bristol, Bristol County, BS1, England, UK
Listed on 2026-07-23
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, IT Project Manager
Location: Bristol or Edinburgh (Hybrid)
In this position, you’ll be based in the Bristol or Edinburgh office for a minimum of three days a week, with flexibility to work from home some days.
OverviewComputershare has an opportunity for a Head of Site Reliability Engineering (SRE) to join our global technology team during a key focus on transforming our organisation toward an SRE operating model.
Reporting directly to the Global Head of Technology Operations, you will operate within Technology Services with a global mandate to establish and mature Site Reliability Engineering capabilities across the organisation. You will partner closely with Engineering, Infrastructure Operations, and Security to improve service reliability, resilience and performance of critical platforms.
A role you will loveWe are seeking an experienced and visionary Head of SRE to define, lead and evolve our global reliability strategy. This senior leadership role is responsible for driving operational excellence, service reliability, observability, automation and continuous improvement across our technology landscape.
Key Responsibilities- Drive adoption of SRE principles (SLOs, error budgets, toil reduction).
- Establish observability and monitoring standards.
- Lead automation‑first operations.
- Improve incident and problem management maturity.
- Partner with software and infrastructure engineering teams to embed reliability into the product lifecycle.
- Establish SRE governance, standards and operating model.
You will be an experienced SRE leader who combines deep technical expertise with the ability to build high‑performing teams and drive operational transformation at scale.
Proven experience building, leading, and developing Site Reliability Engineering or Production Engineering teams, with a strong understanding of SRE principles including service level objectives, service level indicators, error budgets and toil reduction.
Extensive experience driving automation initiatives, with strong scripting and development capabilities using technologies such as Python, Power Shell, Bash, Terraform and Ansible Automation Platform.
Other key skills:
- Robust knowledge of observability and monitoring practices, and experience implementing and managing platforms such as Dynatrace, Prometheus, Grafana, and Splunk.
- Good understanding of CI/CD tooling and modern software delivery practices, including Jenkins, Git Lab CI, and Azure Dev Ops.
- Background spanning both software engineering and technology operations environments.
- Professional certifications in cloud technologies, Site Reliability Engineering, platform engineering or reliability engineering disciplines.
- Passionate about reliability, resilience, automation and continuous improvement.
- Strategic thinker who can balance long‑term vision with operational delivery.
If you’re a confident leader able to inspire teams, challenge traditional ways of working and drive meaningful change, we’d love to hear from you.
Rewards designed for youFlexible work to help you find the best balance between work and lifestyle.
Health and wellbeing rewards that can be tailored to support you and your family.
Invest in our business by setting aside salary to purchase shares in the company, and you’ll receive a company contribution as well.
Extra rewards ranging from recognition awards and team get‑togethers to helping you invest in your future.
And more. Our community is welcoming and close‑knit, with experienced colleagues ready to help you grow. Visit our careers hub for more information:
#J-18808-LjbffrTo Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: