Principal Site Reliability Engineer
On-siteSeattle, Washington, United States
Job Summary
Design and architect infrastructure and services to ensure reliability and functionality, while forecasting demands and responding to capacity needs. Collaborate with software development teams to build reliable, scalable infrastructures and exercise judgment in data collection to optimize operations. Perform incident response and maintenance tasks, providing comprehensive health and performance reporting alongside support for technology documentation. Identify automation opportunities and leverage advanced knowledge to conduct experiments with new tools, maintaining awareness of site reliability trends. Communicate service information and articulate the potential impact of changes proactively.
Required Qualifications
- Advanced knowledge
- Judgment
- Ability to forecast demands
- Ability to respond to capacity needs
- Ability to collaborate with software development teams
- Ability to develop reliable and scalable infrastructures
- Ability to perform data collection
- Ability to maintain and optimize operations and reliability
- Ability to perform incident response
- Ability to perform maintenance tasks
- Ability to provide comprehensive health and performance reporting
- Ability to identify and recommend opportunities for automation
- Ability to communicate comprehensive information about services
- Ability to proactively anticipate and articulate the potential impact of changes
- Ability to provide comprehensive support for technology
- Ability to document incidents
- Ability to conduct advanced experiments with new tools
- Ability to develop and maintain advanced knowledge of site reliability trends
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.