Site Reliability Engineer 2
On-siteSingapore, Singapore
Job Summary
Contribute to a team building and maintaining workflows that automate releasing, testing, and deploying products while improving staging and production environments for SaaS distributions. Research and build monitoring/analyze tools to optimize code-base building and deployment, manage distributed systems, and automate delivery across AWS, GCP, Azure, and container technologies like Docker and Kubernetes. Design, implement, and orchestrate Kubernetes container clusters while debugging cluster issues and networking problems in a 24/7/365 service environment. This role supports Kong Inc.'s unified API and AI platform, Kong Konnect, which secures and governs intelligence flow across APIs and AI models for Fortune 500 clients and startups.
Required Qualifications
- BS degree in Computer Science, similar technical field of study or equivalent practical experience
- Located in Singpaore
- Experience with continuous/rapid release engineering (CI/CD)
- Infrastructure as Code configuration management systems such as Terraform, Chef, Puppet or Ansible
- Experience with Apache Kafka
- Experience building and administering alerting and monitoring systems for API services
- Strong knowledge of Linux/Unix systems
- Knowledge of one or more mainstream programming languages (Go, C/C++, Python)
- Strong skills in network services such as DNS, TLS/SSL, HTTP
- Experience working in a 24/7/365 service environment
- Debugging Kubernetes clusters, issues, and networking problems
- Design, implement, manage and orchestrate Kubernetes container clusters
Desired Qualifications
- Experience implementing secure and highly available distributed systems/ microservices
- Working knowledge of CSP networking solutions AWS (Transit Gateway, Direct Connect, VPC peering, VPN), Azure (Azure VNet ) and GCP (Network Connectivity Center)
- Professional experience managing production software in AWS
- Experience with PostgreSQL in multi-region configuration
- Experience with Datadog, ElasticSearch, Prometheus, Grafana
- Experience with Redis in multi-region configuration
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.