Site Reliability Engineer (SRE) / DevOps Engineer -
On-siteDublin, Leinster, Ireland
Job Summary
Own and operate Kafka-based messaging platforms in production environments while applying SRE principles to improve reliability, availability, and performance. Drive DevOps and automation initiatives to reduce toil by building enhancements using Ansible, scripts, and CI/CD pipelines. Perform incident management, root cause analysis, capacity planning, and operational readiness for banking, insurance, retail, and healthcare clients. Collaborate closely with application and platform engineering teams to contribute to Java-based tooling and platform enhancements. Requires 4–7 years of experience with Kafka architecture, Linux, distributed systems, and monitoring. Fulcrum Digital provides end-to-end digital transformation services across multiple industries.
Required Qualifications
- 4–7 years of experience working with Kafka / messaging systems
- Strong understanding of Kafka architecture (brokers, topics, partitions, replication)
- Hands-on experience with SRE / DevOps practices
- Proven skills in automation (Ansible, scripting, CI/CD)
- Java development background (ability to debug, enhance, or build platform tools)
- Experience with Linux, distributed systems, monitoring & alerting
- Exposure to incident response, production support, and operational excellence
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.