DevOps Engineer
$125,000–$160,000 year
On-siteSan Diego, California, United States
Job Summary
Design, build, and maintain cloud infrastructure on GCP, focusing on GKE, container orchestration, and secure multi-environment deployments using Terraform and Helm charts. Optimize CI/CD pipelines with GitHub Actions to support reliable deployments across staging, beta, and production environments. Implement robust observability across logging, metrics, tracing, and alerting while driving security best practices including secrets management and zero-trust patterns. Participate in on-call rotations, incident response, and post-incident reviews to continuously improve platform reliability. Partner closely with Backend, AI, and Product teams to support new features, data pipelines, and infrastructure requirements as the platform grows.
Required Qualifications
- 3+ years of experience in DevOps, SRE, Infrastructure Engineering, or a related role in cloud-native environments.
- Strong experience with Terraform for provisioning, versioning, and managing cloud infrastructure.
- Hands-on experience working with GKE or another production-grade Kubernetes environment.
- Experience writing, maintaining, and optimizing Helm charts for application deployments.
- Strong understanding of Docker, containerization, microservices architecture, and the Kubernetes ecosystem.
- Experience building and maintaining CI/CD pipelines, ideally using GitHub Actions.
- Familiarity with GCP services including Cloud Run, Cloud Functions, GCS, Pub/Sub, IAM, Cloud SQL, and Secrets Manager.
- Strong understanding of monitoring and observability tools such as Grafana, Prometheus, OpenTelemetry, or Google Cloud Monitoring.
- Solid understanding of networking fundamentals including DNS, load balancing, VPCs, ingress/egress policies, and service meshes.
- Strong communication skills, ownership mindset, and comfort working in a fast-paced startup environment.
Desired Qualifications
- Experience with security and compliance frameworks such as HIPAA, SOC 2, NIST, or CIS Benchmarks.
- Experience supporting AI/ML infrastructure, GPU workloads, or high-volume data pipelines.
- Experience optimizing cost, performance, and reliability in multi-region or high-availability cloud architectures.
- Familiarity with distributed tracing, log aggregation, and advanced observability patterns.
- Experience with feature flag systems, canary deployments, blue-green deployments, or progressive delivery.
- Experience working in a fast-scaling startup or high-growth engineering organization.
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.