Hackajob Ltd

Hackajob Ltd

Glasgow, Scotland

Lead Site Reliability Engineer

Full-Time£100,000 - 100,000 per yearheuteUnited Kingdom
IT

Job Description

Salary: £100,000 - 100,000 per year

Requirements:
  • Formal training or certification on site reliability engineering concepts and advanced applied experience
  • Experience designing and coding complex problems in public cloud environments such as AWS
  • Deep proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices
  • Fluency in Python and deep knowledge of software applications and technical processes
  • Proficiency in observability, including white-box and black-box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, and Splunk
  • Proficiency and experience with continuous integration and continuous delivery tools such as Jenkins, GitLab, and Terraform
  • Experience with containers and container orchestration such as ECS, Kubernetes, and Docker
  • Experience troubleshooting common networking technologies and issues
  • Ability to identify and solve problems related to complex data structures and algorithms
  • Drive to self-educate, evaluate new technology, teach new programming languages to team members, and collaborate across different levels and stakeholder groups
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows
  • Ability to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails, and ensure outcomes align to resiliency and security expectations
Responsibilities:
  • Demonstrate and champion site reliability culture and practices and exert technical influence throughout our team
  • Lead initiatives to improve the reliability and stability of our teams applications and platforms using data-driven analytics to improve service levels
  • Collaborate with team members to identify comprehensive service level indicators and establish reasonable service level objectives and error budgets with customers
  • Demonstrate a high level of technical expertise within one or more technical domains and proactively identify and solve technology-related bottlenecks
  • Act as the main point of contact during major incidents for our application and identify and solve issues quickly to avoid financial losses
  • Document and share knowledge within our organization via internal forums and communities of practice
  • Use enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements
  • Lead reuse-first adoption of AI-assisted reliability workflows across SDLC and toolchain practices, ensuring traceability, auditability, resiliency, and security controls
  • Take lead on resiliency design reviews
  • Break up complex problems into digestible work for other engineers
  • Act as a technical lead for medium to large-sized products
  • Provide advice and mentoring to other engineers
Technologies:
  • AI
  • AWS
  • Cloud
  • Datadog
  • Docker
  • Dynatrace
  • GitLab
  • Grafana
  • Support
  • Jenkins
  • Kubernetes
  • Marketing
  • Prometheus
  • Python
  • Security
  • Splunk
  • Terraform
  • CI/CD

More:

hackajob is partnering directly with JPMorganChase to hire for this role. We are seeking a Lead Site Reliability Engineer within our Infrastructure Platforms team to help define the future of a globally recognized firm. This role places you in a leadership position where you will advise on technical and business issues, support major incident response, and help drive reliability across our applications and platforms. We are a global leader in financial services, providing strategic advice and products to prominent corporations, governments, wealthy individuals, and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do, and we build trusted, long-term partnerships to help clients achieve their business objectives. We value the diverse talents our people bring to our global workforce and are committed to diversity, inclusion, and reasonable accommodations. Our professionals in Corporate Functions support finance, risk, human resources, marketing, and other essential areas that help set our businesses, clients, customers, and employees up for success.

last updated 35 week of 2026

Interested in this role?

Submit your application now

How to Apply

Ready to apply for this position? Here's what you need:

  • An updated resume highlighting relevant experience
  • A compelling cover letter (if required)
  • Portfolio or work samples (for relevant positions)

About Hackajob Ltd

Hackajob Ltd

Hackajob Ltd

Glasgow

IT

Skills & Technologies

PythonGoRustScalaRailsAWSDockerKubernetesTerraformGitCI/CDJenkins

Inferred from job description

Salary Insight

£100,000

This role

£85,000

UK median

This salary is 18% above the UK median for Lead roles85,000/yr).

Based on 2024–2025 UK technology sector benchmarks

Explore More UK Opportunities

Thousands of tech jobs across the United Kingdom