Engineering Manager - Observability (Hybrid, London)
HybridLondon, England, United Kingdom or Dublin, Leinster, Ireland
Job Summary
Manage a growing engineering team focused on building quality platforms and participating in architecting large-scale distributed tracing and metrics systems for service teams. Lead regular retrospectives, capacity planning, and design sessions while reviewing and approving new architectural proposals from your group. This hybrid role is based in London, Aarhus, or Dublin. Candidates require 8+ years of experience in observability tools like OpenTelemetry and Prometheus, along with expertise in Kubernetes operators and SRE practices. The team leverages an AI-first mindset to accelerate execution on large-scale distributed systems processing trillions of events daily.
Required Qualifications
- Experience in Observability and Tracing (Otel, Tempo, Sentry or similar solutions)
- Experience in Observability and Metrics (Prometheus, Thanos or similar solutions)
- Experience in software development, preferably with building Kubernetes operators using ( Python, Bash or Go )
- Experience with large-scale, business-critical Linux environments
- Experience operating within the cloud, AWS & GCP
- Experience with TDD, CI/CD, Chaos Engineering or similar resilience and reliability practices for infrastructure development
- 8+ years industry experience
Desired Qualifications
- Proven ability to work effectively with both local and remote teams
- Rock solid communication skills, verbal and written
- A combination of confidence and independence, with the prudence to know when to ask for help from the rest of the team
- Contributions and involvement in OSS projects
- Experience with being on-call
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.