Principal Site Reliability Engineer
$220,000–$280,000 year
RemotePoland
Job Summary
Partner with software development teams to build reliable, scalable, secure, and cloud-native services. Define and improve Service Level Indicators and Objectives that connect platform reliability to customer impact. Establish best practices for availability, reliability, scalability, performance, observability, and incident response. Design deployment pipelines and release practices that help teams ship safely, while integrating security and reliability into the software development lifecycle. Lead production incident response, root cause analysis, and post-incident learning to drive reliability improvements. Mentor engineers and foster a culture of continuous improvement and operational excellence. This role requires 8+ years of hands-on experience in Site Reliability Engineering, cloud infrastructure, and platform engineering, with a rotating on-call schedule across multiple time zones. Compensation ranges from zł220,000 to zł280,000 PLN annually.
Required Qualifications
- 8+ years of hands-on experience in Site Reliability Engineering, DevOps, cloud infrastructure, platform engineering, software engineering, or related technical discipline
- Demonstrated experience improving reliability, scalability, availability, deployment safety, and operational performance across production systems
- Experience influencing architecture, technical standards, and reliability practices across multiple engineering teams
- Strong programming ability in at least one language such as Go, Python, Node.js, or Java
- Experience operating cloud infrastructure in AWS environments, including services such as ECS, Fargate, and Lambda
- Hands-on experience with infrastructure as code, preferably Terraform, and modern CI/CD tools such as GitHub Actions, CircleCI, Jenkins, or similar
- Experience with observability, monitoring, and incident management tools such as Honeycomb.io, New Relic, CloudWatch, Grafana, or similar
- Working knowledge of relational and NoSQL databases such as Aurora, MySQL, PostgreSQL, MongoDB, DynamoDB, Redshift, SQL Server, or similar
- Ability to use AI-assisted engineering tools responsibly to improve development workflows, debugging, documentation, operational efficiency, and system design support while maintaining accountability for quality, security, and reliability
- Strong collaboration skills with the ability to influence engineering teams, communicate technical tradeoffs clearly, and guide improvements across a distributed organization
- Willingness to participate in a rotating on-call schedule and support teams across multiple time zones when needed
Desired Qualifications
- Experience selecting, evaluating, or operating agentic services or AI-enabled tools for secure and reliable production use
- Experience building scalable systems using SaaS services, APIs, automation, and service-oriented architectures
- Experience improving developer productivity through platform tooling, paved roads, self-service workflows, or automation
- Experience with security-minded engineering practices, including secure SDLC, secrets management, least-privilege access, or compliance-oriented operational controls
- Experience mentoring engineers or leading cross-team reliability initiatives without formal people management authority
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.