Autodoc logo
AutodocPosted 3 weeks ago

SRE Engineer (Cloud Infrastructure) m/f/d

HybridChisinau, Chișinău Municipality, Moldova

Full TimeLarge

Job Summary

Act as the primary point of contact for developers within your domain, handling service-related queries and managing SRE-specific tasks for over 200 services running in GCP/GKE. Maintain and improve current cloud infrastructure to ensure high availability and scalability, while integrating SRE/DevOps best practices into the development lifecycle from architecture planning to deployment. Research, develop, and implement new infrastructure tools through Proof of Concept projects, and partner with the Automation team to build efficient CI/CD pipelines and custom automated workflows. Participate in developing quality metrics (SLIs/SLOs) and maintain comprehensive project documentation, joining the on-call rotation to ensure 24/7 stability of mission-critical services.

Required Qualifications

  • 3+ years as a SRE/DevOps Engineer
  • Proven experience with containerization and orchestration tools
  • Kubernetes is the must
  • Knowledge of SRE/DevOps methodologies, such as CI/CD, IaC, gitOps, etc.
  • Knowledge of at least one tool from the gitOps approach
  • Experience in Cloud based infrastructures
  • Research and troubleshooting skills
  • Experience in administering and tuning relational and columnar databases, specifically PostgreSQL, MySQL, and ClickHouse
  • Experience in deployment and maintenance of distributed high-load systems
  • Experience in development of fault-tolerance mechanisms - clustering, replication, scaling approaches, etc.
  • Configuration of monitoring solutions (Grafana, VictoriaMetric (operator))
  • Good scripting skills (bash / python)
  • Spoken English
  • The position is available for candidates based in Portugal, Poland, the Czech Republic, Moldova or Kazakhstan

Desired Qualifications

  • GKE is preferred
  • FluxCD is preferred
  • Experience in Cloud based infrastructures (GCP is preferred)
  • Configuration of logging/tracing solutions (open telemetry stack, ViktoriaLogs, Grafana Loki, Grafana Tempo are preferred)
  • Hands-on knowledge of maintaining and scaling Elasticsearch
  • Proficiency with message brokers and event-streaming platforms such as Kafka and RabbitMQ
  • Proficiency with GitlabCI
  • Proficiency in developing, maintaining, and refactoring complex Helm charts
  • Experience in migrating applications to Kubernetes
  • Deep understanding of Linux-like OS processes
  • Experience in implementing security controls in containerized environments
  • Boundless desire to automate any processes with an emphasis on improving security
  • Excellent communication skills

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce