Senior Infrastructure Engineer (m/f/d)
On-siteBerlin, State of Berlin, Germany
Job Summary
Design and operate large-scale distributed systems and data platforms for tens of billions of daily events, ensuring reliability, scalability, and cost-efficiency. Drive OneDT platform standardization by migrating to GitOps/ArgoCD, unifying CI/CD pipelines, and establishing a clear developer-ownership model. Build and maintain Kubernetes-based infrastructure, including Spark-on-K8s and Databricks ecosystems, while implementing MLOps capabilities for training, deployment, and serving. Champion observability and automation initiatives to improve reliability and reduce spend through on-call rotations and incident response. Lead blameless postmortems and prevent recurrence via platform improvements. Requires 8+ years in infrastructure engineering with deep cloud and Kubernetes experience.
Required Qualifications
- 8+ years in infrastructure, platform, or back-end engineering, with a track record of building robust distributed systems
- Deep experience with a major cloud provider (AWS or GCP)
- Proficiency in Go, Java, Python, or Scala
- Strong hands-on experience building and operating Kubernetes infrastructure stacks
- Experience with infrastructure-as-code, CI/CD, GitOps, and modern observability tooling
- Operational maturity: you've owned production systems, run on-call, and improved reliability systematically
- Familiarity with data and/or ML infrastructure: Spark, Kafka, data lakes, Databricks, or comparable technologies
Desired Qualifications
- Experience leading large migrations or platform-standardization programs across multiple teams
- Background in high-throughput, low-latency systems such as real-time bidding or event streaming
- Experience with cost governance (FinOps) at scale
- Azure a plus
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.