Staff Infrastructure Engineer
$287,000–$485,000 year
On-siteSan Francisco, California, United States
Job Summary
Run multi-cluster Kubernetes deployments across AWS/GCP/Azure with failover and disaster recovery, building internal tooling to automate dev-to-prod consistency. Design workload isolation strategies balancing cost, performance, and reliability while enforcing RBAC, secrets management, and audit trails. Define enforceable SLOs, manage CI/CD pipelines with Pulumi and GitHub Actions, and maintain Docker environments across all stages. Debug end-to-end issues from APIs to frontend builds, lead incident response, and write actionable postmortems. Own uptime through proactive observability and automate repetitive toil to accelerate team velocity.
Required Qualifications
- 8+ years of experience
- You've actually lived in Kubernetes, not just deployed to it
- You get cluster architecture, scheduling, networking, storage primitives, and the fun ways distributed systems fail
- IaC and CI/CD are second nature
- Pulumi or Terraform
- Docker
- GitHub Actions
- Experience across multi-cluster, multi-region setups, not just a single happy-path environment
- You think in failure modes
- You can turn 'the contract says X' into 'the infra does X'
- You're a DevOps/infra person at your core
- Enough full-stack range to not get stuck when the bug crosses into backend or frontend territory
- You debug like a detective
- Systematic, relentless, not afraid to go five layers deep to find the real root cause
- Ambiguity doesn't scare you
- You trace it across systems until you find it
Desired Qualifications
- Experience working in a startup environment
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.