Senior DevOps engineer
RemoteSingapore, Singapore
Job Summary
Maintain and optimize core server infrastructure including bare-metal servers, LXC containers, virtual machines, and cloud environments. Operate and support services such as Nginx, Puppet, GitLab, and Grafana while managing infrastructure evolution through Infrastructure as Code paradigms. Handle and resolve incidents related to infrastructure operations and design high-availability, fault-tolerant software solutions. Monitor service performance using observability tools to ensure system reliability and optimal resource utilization. Collaborate closely with cross-functional teams including network engineers and developers.
Required Qualifications
- 4+ years of experience in Linux administration, DevOps, or Site Reliability Engineering (SRE)
- Strong proficiency in automating tasks using Bash or similar scripting languages
- Solid understanding of networking fundamentals (TCP/IP stack, routing, DNS, etc.)
- Hands-on experience managing bare-metal infrastructure in production environments
- Experience with configuration management systems such as Ansible or Puppet
- Experience with distributed databases (Elasticsearch, Cassandra, MongoDB, MySQL, PostgreSQL, etc.)
- Expertise in Kubernetes administration and managing containerized workloads
- Experience with IaC tools such as FluxCD or ArgoCD
- Ability to design and implement high-performance, fault-tolerant, and secure infrastructure solutions
- Experience with monitoring and observability systems — Zabbix, VictoriaMetrics, Loki, Grafana, etc. — including building dashboards and configuring alerting
- Programming experience in Python, Go
- Experience with cloud providers (AWS, GCP, Alibaba Cloud, or others)
- Familiarity with distributed storage systems (Ceph)
- Experience with service meshes (e.g., Istio)
- Proven track record of working on high-load, large-scale distributed systems
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.