Site Reliability Engineer Jobs

Engineers who ensure cloud services are reliable, scalable, and efficient. A critical role in maintaining uptime and performance in cloud environments.

Open roles
10
Salary range
£45k – £100k
Hiring companies
6

Site Reliability Engineers (SREs) are the backbone of cloud operations, ensuring that applications and services run smoothly and efficiently. They work closely with development teams to build and maintain robust, scalable systems that can handle high traffic and complex workloads. SREs are in high demand across a range of industries, from tech startups to large enterprises, and their role is crucial in maintaining the reliability and performance of cloud-based applications.

What the role does

Inside the role of a Site Reliability Engineer

A typical week for an SRE is a mix of proactive system maintenance, incident response, and collaboration with development teams.

  1. 01
    Monitor system performance and health metrics.
  2. 02
    Respond to and resolve incidents and outages.
  3. 03
    Implement and optimise automation scripts and tools.
  4. 04
    Collaborate with developers on system design and improvements.
  5. 05
    Conduct post-incident reviews and document findings.
  6. 06
    Participate in on-call rotations and provide 24/7 support when needed.
Salary on the board

£45k – £100k

Based on advertised midpoints across the 66 priced listings posted in the last 12 months. Base salary only.

Salary visibility
38% of listings advertise a salary.
By seniority
£k base
Mid
44
92
19 jobs
Senior
60
85
20 jobs
Lead
64
100
14 jobs
Skills & tools

What hiring managers ask for

% of 58 listings posted in the last 12 months that mention each skill, extracted from job descriptions.

Terraform
67%
CI/CD
64%
Python
53%
AWS
52%
Linux
45%
Observability
41%
Azure
40%
Kubernetes
38%
Grafana
36%
Prometheus
36%
Automation
34%
Monitoring
33%
Career ladder

From Junior to Principal

A typical UK progression for site reliability engineers. Years are guidance — strong people move faster, and many senior folks sidestep into research, product or management.

  1. Level 1

    Junior Site Reliability Engineer

    0–2 yrs

    Assist in monitoring and maintaining cloud infrastructure, with a focus on learning and supporting more experienced team members.

  2. Level 2

    Site Reliability Engineer

    2–5 yrs

    Own the reliability and performance of specific systems, implementing automation and optimisation strategies.

  3. Level 3

    Senior Site Reliability Engineer

    5–8 yrs

    Lead the design and implementation of complex cloud architectures, mentor junior engineers, and drive reliability initiatives.

  4. Level 4

    Principal Site Reliability Engineer

    8+ yrs

    Strategise and oversee the reliability and scalability of the entire cloud infrastructure, influencing company-wide practices and policies.

Pathway

How to become a Site Reliability Engineer

There's no single route, but most people follow some version of these steps.

  1. 1

    Learn the Basics

    Start with foundational knowledge in cloud platforms, scripting, and system administration.

  2. 2

    Gain Practical Experience

    Work on real-world projects, contributing to monitoring, automation, and incident response.

  3. 3

    Specialise in SRE Practices

    Deepen your expertise in reliability engineering, including capacity planning and disaster recovery.

  4. 4

    Lead Projects and Teams

    Take on leadership roles, managing teams and driving large-scale reliability initiatives.

  5. 5

    Influence Company Strategy

    Shape the company's cloud strategy and best practices, contributing to long-term reliability and efficiency.

Live jobs

10 live roles

See all 10 roles

Site Reliability Engineer

Site Reliability Engineer (SRE) - 100% RemoteLocation: Fully RemoteDuration: PermanentAre you passionate about building unbreakable systems and automating away the noise?We are looking for a dedicated Site Reliability Engineer (SRE) to join our remote team. Your primary mission will be...

Randstad Technologies Recruitment London, United Kingdom £50,000 – £60,000 pa
Amazon logo

Site Reliability Engineer, Region Services

Would you like to help implement innovative cloud computing solutions and solve the most complex technical problems? Are you excited by the prospect of building and running the world's largest cloud computing infrastructure to provide a better world for future...

Amazon London, United Kingdom
Permanent
Adecco logo

Site Reliability Engineer (SRE) / Platform Engineer

This role involves designing, building, and securing cloud-native platforms on GCP with a focus on Kubernetes, Terraform, and Python-based automation. The engineer will ensure high availability, reliability, and security of production systems supporting AI-driven applications, while improving CI/CD pipelines and operational tooling. It's a hands-on position emphasizing automation, incident response, and collaboration with development teams in a hybrid work environment.

Adecco London, City And County Of the City Of London, United Kingdom
Hybrid Contract
Amazon logo

Site Reliability Engineer , Cryptography, Access and Identity Services

This role involves building and operating high-availability Cryptography, Access and Identity services within AWS's large-scale cloud infrastructure. The engineer will focus on automation, operational excellence, and reducing manual toil through AI tooling and custom solutions. Responsibilities include troubleshooting Linux and networking issues, driving process improvements, and mentoring junior team members in a collaborative, innovation-driven environment.

Amazon London, United Kingdom
Hybrid Permanent

Site Reliability Engineer (AWS)

This role involves ensuring the reliability and resilience of large-scale cloud platforms running on AWS, with a focus on automation, incident response, and operational excellence. You'll work across the full incident lifecycle, from real-time troubleshooting to post-mortem analysis and systemic improvements. The role supports containerised workloads using Kubernetes and emphasizes engineering solutions over traditional operations, within a 24/7 shift pattern for critical national services.

Spectrum IT Recruitment Southampton, Hampshire, SO19 8NJ, United Kingdom £60,000 pa

Site Reliability Engineer (SRE)

This role involves designing and scaling observability systems to ensure global service reliability and performance. You'll manage large-scale Prometheus, Elasticsearch, and Kafka infrastructure, build alerting workflows, and support self-service tooling for engineering teams. The position requires deep expertise in distributed systems, automation, and on-call incident response.

Randstad Technologies Recruitment United Kingdom £55 – £60 ph
Hybrid Permanent

Site Reliability Engineer - SRE

This role involves improving reliability and efficiency of a market risk platform through automation, AI-driven process optimization, and operational excellence. The engineer will focus on reducing toil, enhancing monitoring, and streamlining workflows in a hybrid on-prem and cloud environment. Work includes deep troubleshooting across distributed systems, implementing guardrails for AI tools, and driving measurable improvements in incident response and system stability.

PRACYVA London, United Kingdom £450 – £500 pd

Site Reliability Engineer - SC Cleared

This role involves working on high-impact, mission-critical projects for Defence and National Security clients. You will be responsible for ensuring the reliability and performance of complex systems, using your expertise in software development, cloud infrastructure, and system monitoring.

Searchability NS&D Gloucestershire, United Kingdom £40,000 – £65,000 pa
Hybrid Permanent Clearance Required
Top hirers

Companies hiring site reliability engineers

See all companies →
Hiring locations

Where this role is hiring

The locations with the most live listings for this role today.

FAQs

Common questions

  • Essential skills include strong knowledge of cloud platforms, scripting, system administration, and automation tools. Familiarity with monitoring and incident response is also crucial.

  • SREs collaborate closely with developers to ensure that applications are designed for reliability and scalability. They provide feedback on system design and help implement automation and monitoring solutions.

  • SREs often work in fast-paced, collaborative environments. They may be part of on-call rotations and need to be available to respond to incidents at any time.

  • Advancement involves gaining experience, specialising in advanced SRE practices, and taking on leadership roles. Continuous learning and staying updated with the latest cloud technologies are also important.

  • Salary ranges can vary widely based on experience, location, and company size. For more detailed information, please refer to the salary section on this page.

Hiring site reliability engineers?

Post your role in 90 seconds and reach the specialist audience that already reads this page.