SolarWinds logo
SolarWindsPosted 1 month ago

Senior Site Reliability Engineer

HybridBrno, South Moravian, Czechia

Part TimeSenior LevelLarge

Job Summary

Senior Site Reliability Engineer at SolarWinds responsible for ownership of production infrastructure, including Kubernetes clusters and data platforms. Leads zero-downtime Kubernetes version upgrades across large-scale environments, ensures reliability, observability, and scalability of data pipelines end-to-end, and manages/scales the ELK stack for centralized logging. Operates and optimizes ClickHouse clusters for high-throughput analytic workloads and administers Kafka clusters with tuning, scaling, and fault-tolerant delivery. Participates in and improves on-call rotations, runbooks, and incident response, drives automation across infrastructure provisioning and deployments, and collaborates with engineering, data, and product teams to embed reliability from day one.

Required Qualifications

  • 8+ years of experience in SRE, DevOps, or infrastructure engineering
  • Hands-on experience with Kubernetes including version upgrade planning, node pool migrations, and zero-downtime rollouts
  • Strong experience with Kafka — operations, tuning, and scaling in production
  • Proficiency with ELK stack (Elasticsearch, Logstash, Kibana) for log management and observability
  • Experience operating ClickHouse or similar columnar databases at scale
  • Solid background in building and maintaining data pipelines reliably in production
  • Infrastructure as Code with Terraform — modules, state management, and multi-environment setups
  • CI/CD expertise using Flux, Jenkins, or Spinnaker
  • Strong Git practices — branching strategies, GitOps workflows
  • Proficiency in at least one programming/scripting language (Python, Go, Bash)
  • Proven on-call experience with a track record of improving alert quality and reducing MTTR
  • Strong automation mindset — eliminate toil, build durable solutions

Desired Qualifications

  • Extreme ownership — you don't wait to be asked, you drive problems to resolution
  • Collaborative team player who lifts those around them
  • Clear communicator across engineering and non-engineering stakeholders
  • 25 days of vacation per year
  • 3 sick days per year
  • 10 study days per year
  • 2 volunteering days per year
  • 4 weeks’ holidays after 5-year tenure, Sabbatical Leave
  • Up to 48 300CZK personal education budget per year
  • Pension or life insurance matching donation up to 3% of the salary or 4000 CZK per month
  • Cash allowance for meals of 95 CZK per working day
  • Unlimited access to LinkedIn Learning
  • English/Czech classes
  • Multisport card
  • Solarian Referral Program

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce