Senior DevOps Engineer & Site Reliability Engineer (SRE)
RemoteUnited Kingdom
Job Summary
Design, automate, and manage cloud infrastructure across AWS, Azure, and GCP using Terraform, CloudFormation, and Bicep. Manage Kubernetes, Docker, and containerized applications while building and maintaining CI/CD pipelines. Implement monitoring, logging, and observability using Prometheus, Grafana, and ELK stacks. Handle production incidents, root cause analysis, and performance optimization to ensure high availability. Enforce cloud security, vulnerability scanning, and SRE practices including SLA, SLO, and disaster recovery protocols. Requires 10+ years of experience with strong troubleshooting skills and comfort working in UK shifts to support business hours.
Required Qualifications
- 10+ years of relevant DevOps/SRE experience
- Strong hands-on AWS and Kubernetes experience
- Excellent troubleshooting, automation, and communication skills
- Comfortable working in UK shifts and supporting UK business hours as required
Desired Qualifications
- Experience working with international/UK clients is preferred
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.