Latest site reliability engineer Jobs

Site Reliability Engineer

This role involves designing, building, and maintaining resilient and scalable infrastructure, automating processes, and improving monitoring and logging systems. You will work closely with engineers and stakeholders to solve complex challenges and contribute to the continuous improvement of technical practices.

Twinstream Limited Montpellier, Gloucestershire, GL50 1SD, United Kingdom £75,000 – £95,000 pa

Site Reliability Engineer

This role involves ensuring the reliability, scalability, and performance of large-scale production platforms through proactive monitoring, incident response, and automation. The engineer will work closely with development and DevOps teams to improve system resilience, implement Infrastructure as Code, and support global platform scalability. Key responsibilities include on-call incident management, root cause analysis, and driving improvements in availability and security using Kubernetes and Azure.

Bristow Holland Ltd London, United Kingdom £55,000 – £60,000 pa

Site Reliability Engineer

This role involves ensuring the reliability, scalability, and performance of large-scale production systems through proactive monitoring, incident response, and automation. The engineer will use Kubernetes and Azure to build and manage cloud infrastructure, implement Infrastructure as Code with Terraform, and collaborate closely with development teams to balance feature delivery with system stability. Participation in an on-call rota supports rapid response to production issues, while ongoing improvements focus on reducing manual work and enhancing platform resilience.

Bristow Holland Ltd Edinburgh, Alba / Scotland, United Kingdom £55,000 – £60,000 pa
CrowdStrike logo

Site Reliability Engineer II , London)

This role involves ensuring the reliability and performance of CrowdStrike's large-scale distributed systems, with a focus on automation, incident response, and operational excellence. The engineer will work across the full stack, troubleshooting hardware and software issues, optimizing system performance, and contributing to platform availability and capacity planning. A strong emphasis is placed on learning, innovation, and using AI to enhance operational workflows.

CrowdStrike United Kingdom
Remote

Site Reliability Engineer - SC Cleared

This role involves working on high-impact, mission-critical projects for Defence and National Security clients. You will be responsible for ensuring the reliability and performance of complex systems, using your expertise in software development, cloud infrastructure, and system monitoring.

Searchability NS&D Gloucestershire, United Kingdom £40,000 – £65,000 pa
Hybrid Permanent Clearance Required

Site Reliability / Software Engineer - SC Cleared

This role involves developing and maintaining secure, scalable software systems while also ensuring high reliability through SRE and DevOps practices. You'll work across the full software lifecycle, contributing to backend services in Java and Python, frontend components in JavaScript/TypeScript, and automating CI/CD and operational workflows. The position supports mission-critical systems in a regulated environment, with a strong focus on monitoring, observability, and performance optimization.

Searchability NS&D Gloucestershire, United Kingdom £45,000 – £65,000 pa
Hybrid Permanent Clearance Required

Senior Site Reliability Engineer

This role involves ensuring the reliability, scalability, and performance of critical banking systems through hands-on SRE engineering and technical leadership. The engineer will build automated, observable infrastructure, lead incident resolution and root cause analysis, and drive SRE best practices across teams. Emphasis is placed on cloud platforms, infrastructure-as-code, and continuous optimisation within a hybrid working environment.

GCS Glasgow, City Of Glasgow, G2 1AL, United Kingdom £75,000 – £95,000 pa

Senior Site Reliability Engineer

Senior Site Reliability Engineer (AWS CDK)Remote UK£65,000 - £75,000 | Sponsorship available in some circumstancesVIQU have partnered with a leading UK technology organisation undergoing significant investment in its cloud platform and engineering capability. As they continue to scale, they are...

VIQU IT Morley, West Yorkshire, United Kingdom £65,000 – £75,000 pa

Senior Site Reliability Engineer (DevTools)

This role involves maintaining and scaling large-scale developer tools infrastructure, including GitLab, TeamCity, and Artifactory, used across a massive monorepo environment. The engineer will modify and extend open- and closed-source platforms, build self-healing systems, and integrate AI-driven solutions like RAG to improve developer experience. The position emphasizes user-centric problem solving, performance optimization, and deep technical ownership in a high-impact AI cloud environment.

Nebius United Kingdom

Senior Site Reliability Engineer (SRE)

Ensure fault-tolerance, scalability, and continuous operation of cloud infrastructure by solving complex engineering challenges across compute, storage, and networking. Work with cutting-edge technologies to optimize large-scale GPU orchestration and AI/ML deployment. Collaborate in a global, fast-paced environment focused on building a full-stack AI cloud platform.

Nebius United Kingdom

Senior Site Reliability Engineer

This role involves ensuring the reliability, performance, and scalability of large-scale distributed systems in a cloud and on-premises environment. You'll develop automation tools, manage Kubernetes infrastructure, and work closely with development teams to improve service quality and deployment practices. The focus is on observability, incident management, and optimising platform health using modern DevOps tooling and cloud-native technologies.

Spectrum IT Recruitment Southampton, Hampshire, United Kingdom £60,000 – £70,000 pa

Lead Site Reliability Engineer (Dynatrace)

We're looking for an experienced Site Reliability Engineer / Observability Engineer with deep Dynatrace expertise to join a major technology and platform engineering programme.This is not a role for someone who has simply used Dynatrace dashboards. We're looking for an...

SF Partners United Kingdom £80,000 – £100,000 pa

Senior AWS Site Reliability Engineer

This role involves ensuring high availability and performance of large-scale distributed systems by building automated solutions and driving reliability across cloud and on-prem environments. You'll collaborate with development teams to improve deployment practices, lead incident response, and optimise infrastructure through observability and performance tuning. The focus is on enhancing system resilience, scalability, and delivery speed using modern DevOps practices and cloud-native technologies.

Spectrum IT Recruitment London, United Kingdom £60,000 – £70,000 pa
Hybrid Permanent

AI Platform & Site Reliability Engineering Consultant

This role involves leading and shaping client engagements focused on reliability, resilience, and modern cloud operations. You will define and embed SRE engagement models, establish SLIs, SLOs, and error budgets, and design observability strategies using metrics, logs, and traces. The role also includes reducing toil through automation and delivering SRE capability assessments.

Akkodis London, United Kingdom £88,000 – £96,000 pa

Senior SRE Engineer

This role involves leading observability and reliability initiatives across large-scale AWS environments, with a primary focus on end-to-end Dynatrace implementation and optimisation. The engineer will design, deploy, and integrate Dynatrace for comprehensive monitoring, alerting, and performance analysis while leveraging Infrastructure as Code, Kubernetes, and CI/CD pipelines. The position requires deep technical ownership, cross-team collaboration, and proactive troubleshooting within complex distributed systems.

SF Partners Birmingham, United Kingdom £100,000 – £110,000 pa