Site Reliability Engineer (AWS)
Key details
- Compensation
- $60,000
Job Description
Salary: £60,000 - 60,000 per year
Requirements
- We ideally want experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment.
- We are looking for exposure to Linux systems administration.
- We want experience with AWS cloud infrastructure.
- We are looking for experience with Kubernetes and Docker.
- We want experience in production support and incident management.
- We are looking for scripting experience in Python, Bash or Go.
- We want familiarity with monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch.
- We are looking for networking fundamentals including DNS, TCP/IP and load balancing.
- We value a passion for automation, continuous improvement and operational excellence.
- Experience with Infrastructure as Code such as Terraform, SRE principles such as SLIs and SLOs, or regulated environments would be beneficial but is not essential.
Responsibilities
- We monitor and maintain highly available production platforms running in AWS.
- We respond to and manage production incidents across a 24/7 service.
- We investigate complex technical issues and restore services quickly and effectively.
- We develop automation to reduce manual operational tasks and improve platform resilience.
- We build and improve monitoring, alerting and observability across cloud environments.
- We work alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence.
- We contribute to post-incident reviews and drive continuous service improvements.
- We support containerised workloads using Kubernetes and Docker.
Technologies
- AI
- AWS
- Bash
- Cloud
- CloudWatch
- Datadog
- Docker
- Grafana
- Incident Management
- Support
- Kubernetes
- Linux
- Load Balancing
- Prometheus
- Python
- Security
- Splunk
- TCP/IP
- Terraform
- DevOps
More
We are a global leader in AI-powered customer experience and cloud technology, expanding our engineering teams following the award of a major government programme. We are building and supporting highly secure, cloud-native platforms that deliver sensitive communication services. This is a fully remote role in the UK on a 24/7 shift pattern across a 28-day rota including days and nights, with a competitive salary, bonus and excellent benefits. You will join an engineering-led organisation where reliability, automation and continuous improvement are central to the platform, and you will work as part of a collaborative SRE team focused on resilient cloud services.
last updated 30 week of 2026
Company & context
Evidence is labeled so you can tell internal community data from public sources.
Range from 95 indexed roles at this employer: $24,000 - $110,000(mid ~73232)
Context may refresh in the background.
Trust-check this listing
Verify scam risk and ghost-job signals before you apply.
Related roles
Browse more remote DevOps Engineer jobs, or every remote job category.
Automation Software Engineer
Automation Software Engineer
Platform Automation Software Engineer
Source: DevITJobs • Last updated 2d ago