Staff DevOps Engineer
RemoteIndia
Job Summary
Lead infrastructure design discussions for new features and platform initiatives while operating in the EST time zone. Review and improve Terraform modules and infrastructure patterns, then architect and optimize CI/CD pipelines to balance speed, reliability, and security. Monitor production systems, identify systemic risks, and proactively eliminate reliability bottlenecks by troubleshooting complex issues across networking, compute, containers, and databases. Lead incident response efforts and conduct structured root cause analyses to improve MTTR and system resilience. Define and refine observability standards, work closely with application engineers to improve deployment workflows, and implement IAM policies and network isolation practices. Drive automation to reduce manual operational toil and mentor engineers to raise overall DevOps maturity across the organization. Contribute to the long-term platform roadmap and infrastructure strategy while evaluating trade-offs between performance, scalability, cost, and security.
Required Qualifications
- 9-14+ years of hands-on experience in DevOps, SRE, or Cloud Engineering
- Deep production experience in AWS
- Advanced expertise in ECS or Kubernetes, containerized microservices architectures
- Experience designing and operating infrastructure for high-scale SaaS environments with strong availability and performance requirements
- Proven experience designing and operating multi-region cloud architectures to enable disaster recovery, high availability, and globally resilient platforms
- Strong hands-on experience with Terraform
- Strong Linux, networking, and distributed systems fundamentals
- Strong understanding of cloud security design and enforcement
- Proven ability to resolve complex production incidents
- Strong systems thinking and architectural trade-off evaluation
- High ownership mindset and cross-team influence capability
- Excellent communication and technical leadership skills
- Expertise in AWS, Docker, Terraform, CI/CD tooling, scripting
- Strong understanding of distributed systems and SaaS scaling patterns
- Experience defining platform standards and best practices
- Ability to influence engineering roadmaps and drive execution
- Must be able to operate in the EST time zone
Desired Qualifications
- Azure/GCP exposure
- Solid CI/CD experience (GitHub Actions preferred)
- Experience supporting Rails production systems
- Familiarity with SOC 2, GDPR, or compliance audits
- Experience with observability tooling (Prometheus, Grafana, OpenTelemetry)
- Experience in SaaS or cybersecurity environments
- Exposure to FinOps practices
- MLOps or AIOps exposure
- Ruby-on-Rails
- Python
- Go
- Bash
- PostgreSQL
- Redis/Valkey
- MongoDB
- DynamoDB
- CloudWatch
- New Relic
- SonarCloud
- Snyk
- container/IaC scanners
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.