Site Reliability Engineer III
On-siteIrvine, California, United States
Job Summary
Guide peers in adopting site reliability engineering best practices and design level architectures. Collaborate with software teams to implement automated CI/CD pipelines and configure infrastructure as code for applications. Leverage enterprise-authorized AI to accelerate incident triage, identify reliability risks, and prioritize recurring toil improvements. Resolve complex problems using service level indicators before they impact customers and apply observability principles for 24x7 operations. Participate in a 24x7 on-call rotation for incident response and production maintenance. This role supports the Global Payments team at JPMorgan Chase by modernizing mission-critical systems through code and cloud infrastructure.
Required Qualifications
- Formal training or certification on site reliability engineering concepts
- 3+ years applied experience
- Proficient in site reliability culture and principles
- Proficient in at least one programming language such as Python, Java/Spring Boot, and .Net
- Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity
- Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements
- Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., Cloud, AI, Android, etc.)
- Experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection
- Experience with continuous integration and continuous delivery tooling
- Familiarity with container and container orchestration and troubleshooting common networking technologies and issues
- Participation in a 24x7 on-call rotation
- This position is subject to Section 19 of the Federal Deposit Insurance Act
- An employment offer for this position is contingent on JPMorgan Chase's review of criminal conviction history, including pretrial diversions or program entries
Desired Qualifications
- Experience operating mixed EKS + EC2 environments, including load balancing and DNS (ALB/NLB, Route 53) and service connectivity troubleshooting
- Experience improving Terraform patterns (modules, reusable configurations) and safe operational practices for state and drift management aligned to team standards
- Familiarity with Kubernetes ingress and certificate deployment patterns (TLS termination, service-to-service encryption/mTLS concepts)
- Experience with observability platforms (CloudWatch plus centralized logging/metrics tooling) and alert tuning for 24x7 operations
- Understanding of controls-focused operations in regulated environments (change management discipline, evidence collection, audit support)
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.