Site Reliability Engineer, SaaS
$145–$175,000 year
HybridNew York City, New York, United States or New York, United States
Job Summary
Own the reliability, performance, and operational health of enterprise SaaS platforms supporting collaboration, automation, and analytics across the organization. Define and track SLIs, SLOs, and error budgets for Microsoft 365, Azure, Slack, Adobe, and Zoom environments while driving incident response and building automation to reduce operational risk. Serve as an escalation point for complex outages, lead root cause analysis, and champion security-first, reliability-first mindsets aligned with SOC2 and ITSM standards. Automate provisioning, onboarding, and request fulfillment using APIs and ITSM tooling to eliminate manual toil. Partner with stakeholders to improve uptime, reduce operational burden, and translate reliability metrics into business impact.
Required Qualifications
- 5+ years of experience in a systems reliability, SRE, or SaaS operations role
- Strong expertise in Microsoft 365 ecosystem (SharePoint, Copilot)
- Azure services or data platforms
- Slack administration and integrations
- Zoom administration
- Email flow
- Experience with identity platforms (Entra ID / SSO / SCIM)
- Strong understanding of SaaS reliability, security, and compliance practices
- Experience with automation and scripting (PowerShell, APIs, scripting)
- Experience defining and monitoring SLAs/SLOs/SLIs
- Systems thinking with strong execution capability
- Incident response and calm-under-pressure decision making
- Cross-platform architecture mindset
- Strong communication and stakeholder management
- Data-driven, metrics-oriented decision making
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.