Hackajob Ltd

Hackajob Ltd

Charing Cross, London

Lead SRE - Charing Cross

Full-Time£61,000 - 101,000 per year2 days agoUnited Kingdom
IT

Job Description

Salary: £61,000 - 101,000 per year

Requirements:
  • Formal training or certification in software engineering concepts, with advanced applied experience.
  • Proven software engineering experience and proficiency in at least one programming language, such as Python, Go, or Java.
  • Experience designing, coding, testing, and delivering software in at least one technology stack.
  • Strong debugging and troubleshooting skills across distributed systems.
  • Experience as a Site Reliability Engineer or supporting production services in an SRE capacity.
  • Working knowledge of microservice infrastructure components, including service discovery, ingress, networking, and load balancing.
  • Experience with Kubernetes and cloud computing services.
  • Familiarity with observability and reliability tools such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger.
  • Ability to use AI-assisted engineering tools responsibly, validate outputs, understand failure modes, and handle sensitive information securely.
  • Preferred: Experience with AWS.
  • Preferred: Experience building internal reliability tooling, such as command-line tools, automation pipelines, operators, or controllers.
  • Preferred: Experience improving developer experience through golden paths, paved roads, templates, or reusable engineering patterns.
  • Preferred: Experience applying AI to operational workflows, such as alert enrichment, summarisation, runbook generation, or anomaly triage, using approved tools and patterns.
Responsibilities:
  • Improve the reliability, monitoring, and alerting of mission-critical microservices.
  • Reduce operational toil by automating processes and building reliable infrastructure and tooling that accelerates feature development.
  • Develop service metrics, user journeys, service-level indicators and objectives, error budgets, dashboards, and actionable alerts.
  • Work with development teams throughout the software lifecycle to design services for reliability and scale.
  • Design and implement self-healing and resiliency patterns, including graceful degradation, rate limiting, circuit breakers, and failover strategies.
  • Partner with engineering, product, and platform teams to promote reliability standards and adoption.
  • Conduct performance testing and capacity planning to identify and address bottlenecks proactively.
  • Participate in feature planning to incorporate metrics, alerting, logging, automation, resiliency, capacity, and performance needs from the outset.
  • Use approved AI tools to support root-cause analysis, log and trace investigation, runbook drafting, post-incident analysis, test scaffolding, and documentation.
  • Develop role-relevant AI skills, including effective prompting, output validation, automation workflows, and safe usage practices.
Technologies:
  • AI
  • AWS
  • Cloud
  • ElasticSearch
  • Grafana
  • Support
  • Java
  • Kibana
  • Kubernetes
  • Load Balancing
  • Prometheus
  • Python
  • microservices
  • Network

More:

We are a global financial services leader serving prominent corporations, governments, wealthy individuals, and institutional investors, with a focus on trusted, long-term client partnerships. Our International Consumer Bank is expanding from the US into the UK and Europe, transforming digital banking through intuitive customer experiences. As a Site Reliability Engineer, you will join a diverse, inclusive, geographically distributed team working to improve the reliability, resilience, observability, and operability of customer-facing digital banking services. Our Corporate Technology team develops applications and provides technology support across functions including Finance, Treasury, Risk Management, Human Resources, Compliance, and Legal, while supporting evolving technology needs and controls.

last updated 40 week of 2026

Interested in this role?

Submit your application now

How to Apply

Ready to apply for this position? Here's what you need:

  • An updated resume highlighting relevant experience
  • A compelling cover letter (if required)
  • Portfolio or work samples (for relevant positions)

About Hackajob Ltd

Hackajob Ltd

Hackajob Ltd

Charing Cross

IT

Skills & Technologies

PythonJavaGoRustAWSKubernetesGitElasticsearchAISREUI

Inferred from job description

Salary Insight

£81,000

This role

£85,000

UK median

This salary is 5% below the UK median for Lead roles (£85,000/yr).

Based on 2024–2025 UK technology sector benchmarks

Explore More UK Opportunities

Thousands of tech jobs across the United Kingdom