Team Manager, SRE
RemoteIndia
Job Summary
Lead and mentor a team of Site Reliability Engineers to ensure technical excellence, timely incident resolution, and professional growth. Drive operational excellence by optimizing workload distribution, automating toil, and meeting strict SLA targets for high-profile enterprise clients. Act as the primary point of contact for critical severity-1 escalations and provide 24/7 management on-call support. Conduct performance reviews, knowledge-sharing sessions, and one-on-ones to strengthen capabilities and disseminate leadership direction. Collaborate with the Director on resource planning and goal setting while serving as a strategic technical advisor to key clients. Foster a culture of psychological safety through proactive problem-solving and blameless post-mortems. Complete an online technical assessment as part of the application process.
Required Qualifications
- A minimum of 3 years of previous experience leading and managing technical teams, specifically with a track record of guiding senior-level SRE and DevOps engineers
- Proven experience managing complex, high-stakes enterprise client relationships
- Ability to confidently interface with high-level client stakeholders and act as a trusted advisor
- Exceptional verbal and written communication skills
- Must be able to seamlessly translate complex technical challenges and strategic goals between C-level executives and highly technical engineering teams
- Demonstrated resilience and composure under pressure
- Strong crisis management and escalation handling skills
- Proven ability to steer teams through significant technical, operational, and client-facing challenges
- Proven experience in SRE and/or DevOps
- Strong SRE and DevOps mindset
- Relentless focus on CI/CD orchestration
- Release management
- Deployment automation
- Scalability
- Toil reduction
- System reliability
- Experience with Google Cloud Platform (GCP)
- Infrastructure as Code (IaC) tools, particularly Terraform
- Strong knowledge of microservices architecture
- Containerization (Kubernetes, Docker)
- Advanced networking concepts
- CI/CD pipeline tools (e.g., Jenkins, GitLab CI, GitHub Actions)
- Hands-on experience with Public Key Infrastructure (PKI)
- Service mesh technologies
- Linux systems administration
- Must be able to lift 50 lbs
Desired Qualifications
- Must be available for weekend shifts
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.