National Health Service

National Health Service

London

Senior Site Reliability Engineer (SRE)

Full-Time£61,000 - 101,000 per yearपरसोंUnited Kingdom
IT

Job Description

Salary: £61,000 - 101,000 per year

Requirements:
  • We require experience as a Site Reliability Engineer, DevOps Engineer, Operations Engineer, or in a similar role.
  • We require coding skills in programming or scripting languages such as Python, PowerShell, or Bash.
  • We require an understanding of Linux/Unix and Windows systems, networking, and distributed systems.
  • We require experience with observability tools such as Prometheus, Grafana, or Datadog, and alerting systems.
  • We require an understanding of infrastructure automation tools such as Terraform, Ansible, PowerShell, or Helm.
  • We require excellent communication and collaboration skills, strong problem-solving ability, and the ability to respond to unexpected demands.
  • Desirable: experience with CI/CD pipelines, cloud platforms such as AWS, GCP, or Azure, and container orchestration such as Kubernetes.
  • Desirable: experience with post-incident reviews, promoting SRE practices across an organization, or training and mentoring junior engineers.
  • We assess candidates on changing and improving, working together, managing a quality service, and delivering at pace; the application requires a supporting statement of up to 1,000 words and a presentation on automating a complex operational process.
  • Successful candidates must pass a basic Disclosure and Barring Service check and meet Security Check clearance requirements. The posting states that candidates would normally have been resident in the UK for the last five years for these checks.
Responsibilities:
  • We expect you to help ensure our services are stable, scalable, performant, and automated.
  • Respond to production incidents, troubleshoot issues, restore services quickly, and conduct root cause analysis and post-incident reviews.
  • Identify system bottlenecks, tune performance, and plan capacity for current and future workloads.
  • Contribute to effective monitoring, alerting, and observability, refining practices to identify issues early and improve response times.
  • Develop automation and tooling to reduce repetitive manual work and operational overhead; write clear, maintainable, well-tested code and use Infrastructure as Code to improve reliability.
  • Contribute to defining, tracking, and improving SLOs, SLIs, and error budgets, and prioritize operational improvements.
  • Promote SRE principles and work with stakeholders to integrate reliability practices into the development lifecycle.
  • Collaborate with software engineering, DevOps, and infrastructure teams to improve deployment and operational workflows and encourage shared responsibility for reliability.
  • Maintain technical documentation, runbooks, and post-incident reports, and provide training and mentorship to engineering teams.
Technologies:
  • AWS
  • Ansible
  • Azure
  • Bash
  • CI/CD
  • Cloud
  • Datadog
  • DevOps
  • GCP
  • Grafana
  • Helm
  • Kubernetes
  • Linux
  • PowerShell
  • Prometheus
  • Python
  • Security
  • Terraform
  • Unix
  • Windows
  • Support
  • LESS

More:

We are the UK Health Security Agency (UKHSA), and we value an inclusive workplace where everyone matters and differences help us develop innovative solutions. We are recruiting a permanent Site Reliability Engineer to join our HPC & SRE engineering team, combining software and systems engineering to build, improve, and operate reliable production systems. The role is available full-time, part-time, as a job share, or with flexible working. We offer hybrid working from our core HQs or scientific campuses in Birmingham, Chilton, Leeds, Liverpool, London, or Porton; the role normally requires at least 60% of contractual hours on site, averaged over a month. Salary is £41,983–£52,113 per year, depending on location, with a market pay supplement of up to £5,000 pro rata, subject to review.

last updated 39 week of 2026

Interested in this role?

Submit your application now

How to Apply

Ready to apply for this position? Here's what you need:

  • An updated resume highlighting relevant experience
  • A compelling cover letter (if required)
  • Portfolio or work samples (for relevant positions)

About National Health Service

National Health Service

National Health Service

London

IT

Skills & Technologies

PythonScalaAWSAzureGCPKubernetesTerraformLinuxCI/CDRESTAIDevOps

Inferred from job description

Salary Insight

£81,000

This role

£75,000

UK median

This salary is 8% above the UK median for Senior roles (£75,000/yr).

Based on 2024–2025 UK technology sector benchmarks

Explore More UK Opportunities

Thousands of tech jobs across the United Kingdom