Senior Site Reliability Engineer, TS Clearance
$166,000–$220,000 year
On-siteCosta Mesa, California, United States
Job Summary
Architect, deploy, and maintain cloud infrastructure across AWS, Azure, and on-premise environments using Kubernetes, Terraform, and infrastructure as code. Collaborate with multi-disciplined teams to define and execute internal and external deployments while promoting SRE best practices in system resilience, performance monitoring, and high availability. Develop and maintain CI/CD pipelines for automated deployment and lead large, focused projects to improve operational capabilities through root cause analysis and scalable tooling. Manage cloud deployments and build strong relationships with internal and external customers to identify technical solutions. Lead the organization in building sustainable mechanisms to deliver at the pace of business scaling.
Required Qualifications
- Holding active U.S. TOP SECRET security clearance
- 6+ years of engineering experience
- Technical expertise and demonstrated performance in one or more of the following areas: networking, cloud technologies, application development and/or cybersecurity
- Deep knowledge of the Kubernetes ecosystem (Docker, Helm, ArgoCD, Terraform)
- Experience with cloud services (AWS/Azure)
- Experience in software languages such as Go, Python, Rust, or C++
- Experience performing data-driven root cause analysis on complex systems
- Demonstrated ability to train peers or customers on the operation of a product
- Computer Science degree or equivalent
Desired Qualifications
- Experience with managing Kubernetes clusters of hundreds of nodes
- Knowledge of performance improvement techniques, metrics and alerting
- Experience with KubeVirt, qemu, virtualization and hypervisor technologies
- Experience with low-level frameworks, Linux and databases
- Excellent written and verbal communication skills
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.