Platform & Reliability Engineer
$66,000โ$90,000 year
On-siteLimassol, Limassol, Cyprus
Job Summary
Design, build, and maintain bare-metal Kubernetes clusters as part of the strategic migration from GCP to self-hosted infrastructure. Plan and execute the migration of services to this new environment while taking ownership of day-to-day operations, including deployments, troubleshooting, and cluster upgrades. Optimize GitOps and CI/CD pipelines to automate routine infrastructure processes and enhance the observability stack using Prometheus, Grafana, and Loki. This role supports Taonga, a mature live game with complex infrastructure, by enabling future growth through rigorous planning and disciplined execution within a focused on-site team in Limassol.
Required Qualifications
- 2+ years' hands-on experience running Kubernetes in production, including deploying clusters from scratch with kubeadm or k3s, troubleshooting, and managing networking and storage
- Experience with Prometheus and Grafana
- Hands-on experience with GitOps practices and tools such as Argo CD or Flux
- Experience with infrastructure as code using Terraform or Pulumi
- Strong Linux systems administration skills
- A solid understanding of networking, including TCP/IP, DNS and load balancing
- Programming experience with Python, Go or TypeScript
Desired Qualifications
- A genuine passion for games
- Experience exploring Kubernetes source code or building custom controllers
- A willingness to use TypeScript for infrastructure as code and automation
- The ability to work independently, make sound technical decisions and take ownership of systems from end to end
- A broad engineering mindset and a willingness to work across different languages, tools and areas of infrastructure
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.