Senior Site Reliability Engineer II
$104,900–$174,700 year
On-siteTexas, United States or Illinois, United States
Job Summary
Support the Business Services DBA team in all infrastructure builds and upgrades across VM and container platforms. Own prioritization of reliability engineering tasks, lead incident management including post-mortems and Root Cause Analysis, and drive disaster recovery planning and production resilience testing. Ensure alignment to SRE frameworks, standards, and operational best practices while supporting infrastructure cost analysis and optimization initiatives. Assist with on-call support and incident response, and contribute to team readiness through training and coaching. This developed professional role focuses on challenging reliability and toil reduction projects within a team partnering with database administration and SRE groups.
Required Qualifications
- Experience working with cloud platforms (such as Amazon Web Services) and Infrastructure as a Service (IaaS)
- Background in DevOps, site reliability engineering practices, or related areas
- Understanding of site reliability principles and their application in improving software development processes
- Familiarity with continuous integration and delivery tools (e.g., Jenkins, GitLab, Terraform)
- Ability to diagnose and resolve networking, performance, and optimization issues in distributed systems
- Basic knowledge of Linux, networking, and storage fundamentals
- Some experience in supporting database environments such as MySQL and PostgreSQL
- Strong communication skills and ability to collaborate within a diverse team
Desired Qualifications
- Experience with monitoring, logging, and alerting tools (such as PMM3,Grafana, Prometheus, or similar) is a plus
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.