Senior Site Reliability Engineer
$140,000–$180,000 year
On-siteLos Angeles, California, United States
Job Summary
Design, implement, and maintain Cloud and On-Premises infrastructure using Infrastructure-as-Code to support the software team and larger Engineering org. Update, deploy, and scale mission-critical services while reducing operational toil through automation and self-service platforms. Collaborate with engineering teams to create highly available, deployable, and reliable products, leading high-impact reliability, security, and resiliency efforts that influence platform strategy. In your first six months, directly shape the software solutions of the first Mega class satellite. Partner with technical teams to design controls, uphold security standards, and proactively mitigate vulnerabilities.
Required Qualifications
- Bachelor's degree in computer science, Information Technology or a STEM discipline or 5+ years of professional experience in Software Engineering, Site Reliability Engineering or DevOps
- Deep experience with cloud platforms (AWS, GCP, or Azure)
- 3+ years of Infrastructure-as-code deployment processes
- Production experience with Kubernetes and container orchestration
- Solid understanding of networking & Linux internals
Desired Qualifications
- Experience managing application deployment workflows
- Experience managing core infrastructure with high availability requirements
- Experience with high-volume manufacturing, aerospace, or other adjacent environments
- Experience writing Terraform, Ansible, or other common Infrastructure-as-Code and owning the full lifecycle of infrastructure deployments
- Experience working with mission-critical and sensitive systems
- Strong programming skills in at least one language (Go, Python, or similar)
- Track record of driving technical strategy and influencing across teams
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.