Paddle logo
PaddlePosted 2 months ago

Site Reliability Engineer

RemotePortugal or Ireland

Full TimeMid LevelMediumFintech Services

Job Summary

Site Reliability Engineer responsible for developing and maintaining automation and tooling to maximise engineering efficiency, driving production reliability, incident response, and disaster recovery. Collaborate with product teams to enable DevOps practices, define production standards as code, and drive cost optimisation across ECS/Fargate, RDS/Aurora, SQS and observability tooling. Emphasizes contribution to monitoring, SLO tracking, performance investigations, and adoption of GitOps practices; supports AI tooling experimentation with appropriate guardrails. Works across AWS ecosystem and distributed systems, with a focus on improving reliability and developer experience for product engineers.

Required Qualifications

  • Software development background
  • Experience shipping and operating production services
  • Strong fundamentals in testing
  • Code review
  • CI/CD
  • Debugging
  • Experience across the AWS ecosystem
  • Partnering with AWS Solution Architects
  • Experience with microservices and distributed systems at scale
  • Experience with monitoring tools (OpenTelemetry, Honeycomb, Grafana)
  • Linux administration
  • Networking knowledge
  • Security-mindedness
  • Collaborative
  • Interest in AI in software development

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce