Senior Staff Cloud Infrastructure Engineer
HybridBengaluru, Karnataka, India or Kuala Lumpur, Kuala Lumpur, Malaysia
Job Summary
Design and evolve secure, scalable cloud infrastructure across AWS and Azure, establishing engineering standards and automation for Terraform-based provisioning. Architect production Kubernetes environments, particularly Amazon EKS, while building reusable platform capabilities that enable self-service application deployment. Lead technical direction for multi-account architectures, networking, security guardrails, and observability, integrating CI/CD pipelines and policy enforcement into reliable workflows. Engineer high-availability systems with automated remediation, collaborating with security stakeholders to embed controls and drive cost optimization. Mentor infrastructure engineers through hands-on collaboration and design reviews to raise team capability.
Required Qualifications
- Extensive hands-on experience designing, building and operating production cloud infrastructure
- Technical depth expected of a Staff, Senior Staff, Principal or equivalent senior individual contributor
- Strong hands-on AWS architecture and production experience
- Strong hands-on Azure capability
- Advanced hands-on Terraform experience
- Strong hands-on capability with Python and/or Go (Golang) for infrastructure automation and tooling
- Strong hands-on capability with Bash/Shell or similar scripting experience
- Strong production Kubernetes experience
- Hands-on experience designing and operating infrastructure CI/CD workflows
- Experience with GitOps technologies such as ArgoCD or Flux
- Experience with Helm
- Hands-on experience configuring enterprise monitoring and observability platforms
- Strong architectural judgement across scalability, availability, resilience, disaster recovery, capacity and production reliability
- Solid infrastructure networking knowledge
- Strong knowledge of cloud security and IAM principles
- Experience using automated testing, static analysis or policy-as-code to enforce infrastructure standards
- Experience with technologies such as service meshes (Istio, Linkerd)
- Experience with alternative IaC approaches such as Pulumi, AWS CDK or CloudFormation
- Experience with FinOps or cloud cost-management tooling
- Experience with hybrid/on-premise infrastructure
- Experience with telecom, CPaaS or similarly regulated/data-residency-sensitive environments
- Experience influencing architecture and engineering standards
- Experience mentoring other engineers and raising technical capability
- Experience remaining deeply hands-on as an individual contributor
- Experience with multi-account / AWS Organizations environments
- Experience with Azure Landing Zone architecture, governance, networking and security
- Experience with Amazon EKS
- Experience with Datadog or similar observability platforms
- Experience with OPA, Sentinel or similar technologies
- Experience with GitHub Actions, GitLab CI, Jenkins, Azure DevOps or similar
- Experience with routing, DNS, load balancing, firewalls, cloud networking and connectivity between cloud and other environments
- Experience with SIEM/SOAR platforms or closely with Security/SOC teams
- Experience with security guardrails and core production services
- Experience with reusable and tested modules, consistent multi-environment patterns and appropriate policy controls
- Experience with internal tooling that connects cloud services, deployment pipelines, observability platforms and operational workflows
- Experience with dashboards, monitors, alerting and integrations
- Experience with automated remediation and self-healing
- Experience with capacity planning, incident resolution and root-cause analysis
- Experience with reducing configuration drift, manual intervention and operational risk
- Experience with performance, reliability, maintainability and cost evaluation
- Experience with platform engineering to make it easier for engineering teams to provision, deploy, test, operate and recover applications
- Experience with cross-functional engineering to understand application requirements and design appropriate infrastructure solutions
- Experience with technical direction, architecture and engineering standards
- Experience with solving complex technical challenges
- Experience with setting technical direction
- Experience with driving greater automation, Infrastructure as Code and engineering maturity
- Experience with designing and implementing modern cloud infrastructure
- Experience with secure, scalable and highly available cloud infrastructure
- Experience with translating business and engineering requirements into robust technical solutions
- Experience with acting as a senior technical authority for Infrastructure
- Experience with owning key architectural decisions
- Experience with establishing engineering standards
- Experience with challenging existing approaches where better technical solutions can be introduced
- Experience with evolving AWS estate
- Experience with establishing and maturing Azure environment
- Experience with developing reusable infrastructure patterns
- Experience with deploying infrastructure reliably and repeatably
- Experience with building reusable infrastructure capabilities
- Experience with designing and improving infrastructure delivery pipelines
- Experience with integrating Infrastructure as Code, version control, automated testing, policy enforcement and deployment automation
- Experience with architecting, operating and improving production Kubernetes environments
- Experience with ensuring container platforms are scalable, resilient, secure and operationally robust
- Experience with applying strong cloud security practices
- Experience with collaborating with security stakeholders
- Experience with driving greater consistency in how infrastructure is designed, provisioned, changed and operated
- Experience with using automation, policy and engineering practices
- Experience with evaluating cloud services, tooling and architectural approaches
- Experience with identifying opportunities to simplify the platform and improve infrastructure efficiency
- Experience with hands-on collaboration
- Experience with technical and design reviews
- Experience with knowledge sharing
- Experience with raising the team's cloud, automation and engineering capability
- Experience with working closely with Software Engineering and other technical teams
- Experience with challenging assumptions constructively
- Experience with designing appropriate infrastructure solutions
- Experience with mostly WFH basis
- Experience with hybrid WFH/WFO model
- Experience with Bangalore – India location
- Experience with KL - Malaysia location
Desired Qualifications
- Experience of multi-account / AWS Organizations environments strongly valued
- Ideally including Azure Landing Zone architecture, governance, networking and security
- Ability to help establish and mature cloud environments rather than only administer existing services
- Experience with reusable modules, remote/state management, multi-environment infrastructure, testing and scalable IaC patterns and standards
- Practical Bash/Shell or similar scripting experience
- Genuine engineering capability beyond occasional or incidental scripting
- A strong track record of engineering repetitive operational activities out of infrastructure through automation, reusable tooling, API integrations, auto-remediation and self-service capabilities
- Ability to architect, troubleshoot and improve containerised production environments
- Experience with ArgoCD or Flux
- Experience with Helm
- Experience with Datadog or similar
- Experience with monitors, dashboards, alerting and third-party integrations rather than only consuming existing dashboards
- Experience with cloud security and IAM principles
- Experience with secure infrastructure design, access controls, hardening and security-conscious automation
- Experience working with SIEM/SOAR platforms or closely with Security/SOC teams would be an advantage
- Experience with OPA, Sentinel or similar technologies would be particularly relevant
- Experience with technologies such as service meshes (Istio, Linkerd), alternative IaC approaches such as Pulumi, AWS CDK or CloudFormation, or comparable modern cloud-platform tooling would be an advantage
- Experience with FinOps or cloud cost-management tooling, hybrid/on-premise infrastructure, and telecom, CPaaS or similarly regulated/data-residency-sensitive environments would be beneficial
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.