PAR logo
PARPosted 1 week ago

Site Reliability Engineer I

HybridGurugram, Haryana, India or Jaipur, Rajasthan, India

Full TimeBachelors DegreeMedium

Job Summary

Monitor and support the reliability, performance, and scalability of PAR's cloud platforms and SaaS services, including building observability solutions and tracking SLIs and SLOs. Participate in incident response, root cause analysis, and blameless postmortems to drive continuous improvement. Develop automation scripts using AI-assisted tools to reduce operational toil and assist with Infrastructure as Code initiatives using Terraform. Support Kubernetes platform operations, capacity planning, and disaster recovery activities while collaborating with Engineering and Security teams on vulnerability remediation. Assist in creating runbooks, operational procedures, and troubleshooting workflows. Participate in rotational shift coverage and on-call support for critical production systems across multiple regions.

Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, Information Technology, or related field, or equivalent practical experience
  • 3+ years of experience supporting cloud-native applications and infrastructure in AWS and GCP pr Azure environments
  • Hands-on experience with Linux systems administration and exposure to Windows systems administration
  • Experience with scripting/automation using Python, Bash, PowerShell, NodeJS, or similar languages
  • Experience with Infrastructure as Code concepts and tools such as Terraform
  • Experience with RDS, MySQL, PostgreSQL, or other managed cloud database services
  • Working understanding of cloud networking, DNS, load balancing, and application performance troubleshooting
  • Experienced in an incident management processes and root cause analysis methodologies
  • Familiarity with CI/CD pipelines and modern software delivery practices
  • Strong troubleshooting, analytical, and problem-solving skills
  • Excellent written and verbal communication skills
  • Ability to write clear documentation – runbooks, operational procedures, postmortems, and knowledge articles
  • Experience supporting production applications in a highly available environment (SaaS environment a plus)
  • Experience with at least one observability/monitoring platform such as Datadog, New Relic, Grafana, Prometheus
  • Working knowledge of containers and Kubernetes-based platforms
  • Experience with GitHub Actions, GitLab CI/CD, Azure DevOps, Jenkins, or similar delivery platforms

Desired Qualifications

  • Experience supporting production applications in a highly available environment (SaaS environment a plus)
  • AWS Certified Cloud Practitioner, Solutions Architect Associate or DevOps Engineer certification
  • Familiarity with AI-assisted development or operational tools (e.g., GitHub Copilot, Amazon Q Developer, AWS Bedrock, or similar)
  • Exposure to building AI agents or intelligent operational assistants
  • Familiarity with AIOps, anomaly detection, predictive monitoring, or intelligent alerting concepts
  • Awareness of Platform Engineering concepts and Internal Developer Platforms (IDP)
  • Exposure to highly regulated or large-scale SaaS environments

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce