DevOps Engineer
$125,000–$145,000 year
HybridSan Diego, California, United States
Job Summary
Build, maintain, and optimize CI/CD pipelines for reliable application deployments while managing AWS infrastructure across ECS, EKS, and EC2 environments. Develop and maintain Infrastructure as Code using Terraform, leveraging AI-assisted tooling to draft, refactor, and review modules and manifests. Monitor system health and performance using Grafana, troubleshoot infrastructure issues in staging and production, and implement logging, alerting, and incident response processes. Collaborate with developers to improve deployment workflows and support containerized workloads using Docker and Kubernetes. Contribute to cost optimization and performance tuning within AWS, maintain clear documentation, and proactively identify reliability, performance, and cost improvements. This full-time, permanent role works out of the San Diego office on a hybrid model (3 days per week in the office).
Required Qualifications
- 2–4 years of experience in DevOps, SRE, or related roles
- Hands-on experience with AWS services, particularly ECS, EKS, and EC2
- Experience writing and maintaining Infrastructure as Code using Terraform
- Experience with monitoring/observability tools, especially Grafana
- Familiarity with CI/CD tools (e.g., GitHub Actions, GitLab CI, Jenkins)
- Experience with containerization (Docker)
- Working knowledge of Linux systems and networking fundamentals
- Scripting experience (Bash, Python, or similar)
- Can independently handle day-to-day infrastructure and deployment tasks
- Comfortable debugging issues across application and infrastructure layers
- Takes ownership of services (e.g., a cluster, pipeline, or environment)
- Knows when to escalate complex architectural decisions
- Proactively identifies reliability, performance, and cost improvements
- Experience applying AI to DevOps practices, including infrastructure automation, CI/CD optimization, incident investigation, observability, and operational efficiency
Desired Qualifications
- Experience operating Kubernetes clusters in production (EKS)
- Familiarity with Prometheus or other metrics backends used with Grafana
- Experience with AWS IAM, networking (VPCs, security groups), and load balancing
- Exposure to autoscaling, high availability, and fault-tolerant system design
- Experience with blue/green or canary deployments
- Understanding of cost monitoring and optimization in AWS
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.