Monitoring Specialist
RemotePoland or Bulgaria
EXPIREDPoland or BulgariaRemoteFull TimeSmall
Full TimeSmall
Job Summary
Monitor infrastructure and application health using Prometheus, Grafana, and Elasticsearch to detect and resolve issues before impacting players. Analyze metrics, logs, and alerts to perform root cause analysis on incidents, optimize monitoring systems by reducing alert noise, and document solutions. Collaborate with development and operations teams during incident response and participate in post-incident reviews. Maintain 24×7 coverage across European time zones in a shift-based schedule.
Required Qualifications
- Solid technical foundation in system administration, DevOps, or technical support roles
- Strong understanding of server, network, and application performance metrics (CPU, memory, latency, RPS)
- Experience with log analysis tools such as Elasticsearch, Kibana, Loki, or Splunk
- Hands-on experience configuring monitoring tools like Prometheus, Alertmanager, or Datadog
- Proficiency with Linux systems and command-line operations
- Flexibility to work shift-based schedules during non-business hours (European time alignment)
- Willingness to work in a shift-based schedule
Desired Qualifications
- Kubernetes hand-on experience
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.