Bank of America logo
Bank of AmericaPosted 1 month ago

Senior Site Reliability Engineer

$152,600–$191,500 year

On-siteCharlotte, North Carolina, United States or Jersey City, New Jersey, United States

Full TimeSenior LevelLarge

Job Summary

Senior GCP Site Reliability Engineer at Bank of America responsible for designing, implementing, and maturing reliability engineering capabilities. Leads reliability patterns for Azure landing zones, private networking, DNS, firewalls, and workload onboarding; drives observability through instrumentation, dashboards, and enterprise monitoring tools; defines SLIs/SLOs and alerting standards; develops reusable Terraform modules and CI/CD patterns; leads incident investigations and post-incident reviews; mentors SRE engineers and collaborates with security/governance teams to integrate IAM and policy-as-code; supports capacity planning and production-readiness for new Azure services and workloads. Requires 4+ years in cloud infrastructure or platform engineering, strong IaC experience with Terraform, familiarity with GCP services, CI/CD pipelines, and DevSecOps practices, plus solid troubleshooting and cross-functional collaboration skills.

Required Qualifications

  • 4+ years of experience in cloud infrastructure engineering, platform engineering, or cloud operations
  • Strong hands-on experience with Infrastructure as Code (IaC), including practical use of Terraform or Terraform Enterprise
  • Solid understanding of software engineering fundamentals, including version control, code quality, and basic testing practices for infrastructure code
  • Experience developing and maintaining Terraform modules and infrastructure configurations
  • Familiarity with CI/CD pipelines for infrastructure deployment, including automated build, test, and release processes
  • Working knowledge of DevSecOps practices, including integrating security and compliance checks into automated workflows
  • Good understanding of GCP services and cloud architecture fundamentals, including networking (VPCs, IAM, load balancing)
  • Exposure to policy-as-code, governance, and compliance requirements in enterprise environments
  • Experience supporting automation and standardization efforts to improve consistency and efficiency in cloud deployments
  • Understanding of monitoring, logging, and observability tools to support system reliability and performance
  • Hands-on experience with incident response, troubleshooting, and root cause analysis in cloud or distributed systems

Desired Qualifications

  • Strong problem-solving and analytical skills
  • Effective communication skills across cross-functional teams
  • Interest in learning and applying emerging technologies and automation techniques (including AI/ML where applicable) to improve platform reliability

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce