Leave us your email address and we'll send you all the new jobs according to your preferences.

Head of Site Reliability Engineering (SRE)

Posted 14 days ago by Computershare

Permanent
Full Time
Other
Gloucestershire, Bristol, United Kingdom, BS153
Job Description

Location: Bristol or Edinburgh (Hybrid)

In this position, you'll be based in the Bristol or Edinburgh office for a minimum of three days a week, with flexibility to work from home some days.

Overview

Computershare has an opportunity for a Head of Site Reliability Engineering (SRE) to join our global technology team during a key focus on transforming our organisation toward an SRE operating model.

Reporting directly to the Global Head of Technology Operations, you will operate within Technology Services with a global mandate to establish and mature Site Reliability Engineering capabilities across the organisation. You will partner closely with Engineering, Infrastructure Operations, and Security to improve service reliability, resilience and performance of critical platforms.

A role you will love

We are seeking an experienced and visionary Head of SRE to define, lead and evolve our global reliability strategy. This senior leadership role is responsible for driving operational excellence, service reliability, observability, automation and continuous improvement across our technology landscape.

Key Responsibilities
  • Drive adoption of SRE principles (SLOs, error budgets, toil reduction).
  • Establish observability and monitoring standards.
  • Lead automation first operations.
  • Improve incident and problem management maturity.
  • Partner with software and infrastructure engineering teams to embed reliability into the product lifecycle.
  • Establish SRE governance, standards and operating model.
What you bring to the role

You will be an experienced SRE leader who combines deep technical expertise with the ability to build high performing teams and drive operational transformation at scale.

Proven experience building, leading, and developing Site Reliability Engineering or Production Engineering teams, with a strong understanding of SRE principles including service level objectives, service level indicators, error budgets and toil reduction.

Extensive experience driving automation initiatives, with strong scripting and development capabilities using technologies such as Python, PowerShell, Bash, Terraform and Ansible Automation Platform.

Other key skills:

  • Robust knowledge of observability and monitoring practices, and experience implementing and managing platforms such as Dynatrace, Prometheus, Grafana, and Splunk.
  • Good understanding of CI/CD tooling and modern software delivery practices, including Jenkins, GitLab CI, and Azure DevOps.
  • Background spanning both software engineering and technology operations environments.
  • Professional certifications in cloud technologies, Site Reliability Engineering, platform engineering or reliability engineering disciplines.
  • Passionate about reliability, resilience, automation and continuous improvement.
  • Strategic thinker who can balance long term vision with operational delivery.
Join us

If you're a confident leader able to inspire teams, challenge traditional ways of working and drive meaningful change, we'd love to hear from you.

Rewards designed for you

Flexible work to help you find the best balance between work and lifestyle.

Health and wellbeing rewards that can be tailored to support you and your family.

Invest in our business by setting aside salary to purchase shares in the company, and you'll receive a company contribution as well.

Extra rewards ranging from recognition awards and team get togethers to helping you invest in your future.

And more. Our community is welcoming and close knit, with experienced colleagues ready to help you grow. Visit our careers hub for more information:

Email this Job