Cloud Platform Engineer
On-siteLondon, England, United Kingdom
Job Summary
Design and operate the AWS cloud infrastructure supporting data platforms, AI/ML services, APIs, and web applications. Develop reusable infrastructure patterns, automation frameworks, and CI/CD workflows using Infrastructure as Code to ensure scalable, secure, and resilient deployments. Implement networking patterns, security controls, and observability tooling to maintain platform health and drive operational excellence across data-driven products. Partner with Data Engineers and Developers to deliver automated solutions that enhance reliability and performance. Investigate complex infrastructure issues, perform root-cause analysis, and identify opportunities to improve cost efficiency and scalability.
Required Qualifications
- Degree in Computer Science, Information Technology, Engineering, or equivalent practical experience
- 5+ years of experience in Platform Engineering, Cloud Engineering, DevOps, Site Reliability Engineering, Data Platform Engineering, or similar roles
- Excellent verbal and written communication skills
- Excellent command of the English language
- Strong hands-on experience designing, building, and operating production workloads on AWS, including serverless, event-driven, and containerized architectures
- Solid understanding of AWS cloud infrastructure, networking, and security, including routing, DNS, load balancing, TLS, network segmentation, VPN connectivity, IAM, encryption, secrets management, vulnerability management, perimeter protection, and security monitoring
- Experience supporting data engineering, analytics, business intelligence, and reporting platforms within AWS environments
- Proficiency with Infrastructure as Code (Terraform, OpenTofu, Ansible, or similar), automation using Python and/or Node.js, and building cloud-native services
- Experience implementing CI/CD pipelines, GitOps workflows, and modern cloud deployment practices, with a strong focus on automation and operational efficiency
- Strong knowledge of observability, monitoring, alerting, incident response, and troubleshooting complex infrastructure, networking, security, and deployment issues
- Strong analytical and problem-solving skills, with the ability to diagnose complex technical issues, take ownership, and drive solutions through to completion
- Self-motivated and proactive, comfortable working in a lean, fast-paced environment, contributing across multiple disciplines with minimal supervision
- Systems-thinking mindset with the ability to design, optimize, and make sound technical decisions for end-to-end platform architectures in ambiguous situations
- Strong security awareness and understanding of operational risk, compliance, governance, and engineering best practices
- Excellent collaboration, communication, and documentation skills, with the ability to work effectively across cross-functional technical and business teams
- Passion for automation, continuous improvement, engineering excellence, and staying current with emerging cloud, platform engineering, data, and AI technologies
- By clicking 'Apply' for this Job, you agree that you have read and accepted our Privacy Statement relating to job applicants and that you provide your consent for the processing of your personal data for the purposes described therein
Desired Qualifications
- Experience supporting machine learning workloads or MLOps practices in cloud environments
- Experience with modern data lake, streaming, analytics, or cloud-native BI architectures
- Understanding of AWS cost optimisation and FinOps principles
- Experience operating in regulated, security-sensitive, or high-availability environments
- Experience with test automation frameworks such as PyTest, UnitTest, Vitest, Jest, or similar
- AWS certifications in Architecture, DevOps, Security, Networking, Data, or Machine Learning
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.