Platform Infrastructure Engineer
On-sitePune, Maharashtra, India
Job Summary
Design, provision, and maintain multi-cloud Kubernetes clusters (EKS, AKS, GKE) using Terraform and Helm across SaaS, BYOC, and on-prem topologies. Own the GitOps delivery pipeline by managing ArgoCD application sets, defining promotion workflows, and enforcing policies via OPA/Gatekeeper. Build and maintain GitHub Actions CI pipelines including build caching, container image signing, vulnerability scanning, and automated integration tests. Define and enforce platform security posture through Kubernetes RBAC, network policies, secret rotation via Vault, and audit logging to SIEM. Implement observability using Prometheus, Grafana, AlertManager, Jaeger/Tempo, and SLO-based alerting. Manage the data platform infrastructure layer for Kafka, PostgreSQL HA, Redis, and Spark jobs. Act as the infrastructure partner for product engineers by reviewing manifests, advising on resource sizing, and building self-service tooling to reduce toil.
Required Qualifications
- 4+ years of infrastructure or platform engineering experience
- at least 2 years operating production Kubernetes clusters at enterprise scale
- Deep Terraform expertise: module authoring, remote state management, workspace strategies for multi-environment/multi-cloud provisioning
- Hands-on ArgoCD or Flux experience for GitOps-based application delivery, including multi-cluster federation and rollback automation
- Strong security engineering background: cloud IAM, Kubernetes RBAC, network policy design, zero-trust networking principles, and compliance frameworks (SOC 2, ISO 27001)
- Scripting and automation proficiency in Python or Go for building internal platform tooling, operator controllers, or custom Kubernetes admission webhooks
- Experience managing stateful workloads on Kubernetes: PostgreSQL, Redis, Kafka — including backup/restore strategies, failover testing, and capacity planning
Desired Qualifications
- Experience delivering air-gap or on-premises enterprise deployments with offline container registries, local Helm chart repositories, and network-restricted environments
- Background with service mesh technologies (Istio, Linkerd) for mTLS, traffic management, and fine-grained observability between Rubiscape's microservices
- Familiarity with FinOps practices — cloud cost allocation, Spot/Preemptible instance strategies, and right-sizing recommendations for Kubernetes workloads
- Kubernetes operator development experience (Kubebuilder, Operator SDK) for automating complex stateful application lifecycle management
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.