Manager, Site Reliability Engineering
On-siteReston, Virginia, United States
Job Summary
Design and architect infrastructure and services for reliability, forecasting demands to ensure adequate resources. Collaborate with software development teams to create scalable infrastructures, advising on data collection and optimizing operations. Lead incident response activities, serving as the escalation point while reviewing health and performance reports. Train team members on automation, change impact communication, and new technology experimentation to build site reliability knowledge. Provide day-to-day direction for improving service functionality and infrastructure reliability.
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.