Lio logo
LioPosted 3 weeks ago

Site Reliability Engineer (SRE) (m/w/d)

On-siteMunich, Bavaria, Germany

Full TimeSmall

Job Summary

Work directly with the CTO on architecture and scaling of infrastructure for production AI workloads, designing multi-region deployments and high-availability systems. Optimize databases for agentic AI, including Vector Search and Hybrid RAG, while improving backend performance and resource efficiency. Build observability, SLOs, alerting, and incident response processes, then automate deployments and developer workflows. Partner closely with product engineering teams to ensure systems scale with rapid growth. This full-time, on-site role in Munich requires strong Python skills and experience with distributed systems and cloud platforms.

Required Qualifications

  • Experience operating production workloads on a major cloud platform
  • Strong Python skills and experience optimizing backend services
  • Solid understanding of distributed systems, asynchronous processing, and scalable architectures
  • Experience with monitoring, observability, and incident management
  • Experience optimizing databases at scale
  • 100% on-site in our Munich office

Desired Qualifications

  • MongoDB
  • Familiarity with CI/CD pipelines
  • GitHub Actions
  • Passion for automation, infrastructure, and solving complex scaling challenges

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce