Platform Operations Engineer - TS/SCI
$107,900–$195,050 year
HybridBethesda, Maryland, United States
Job Summary
Ensure availability, reliability, and performance of a full stack, containerized microservices platform running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and Java. Partner with systems engineers to lead triage, troubleshooting, root cause analysis, and post incident reviews while defining SLIs and SLOs. Participate in release planning, scrums, and cross-team coordination within a mission-focused team supporting the Defense Intelligence Enterprise. Work onsite in Bethesda, MD with a flexible hybrid schedule. Requires active TS/SCI clearance and 8+ years of experience.
Required Qualifications
- BS in Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience)
- 8+ years of relevant experience
- 6+ years with a Master's
- Active TS/SCI clearance with the ability to obtain and maintain a polygraph
- At least one DoD 8570.01 M IAT Level II+ certification (e.g., Security+ CE, CySA+, CCNA Security, SSCP, CISSP (or Associate))
- Ability to obtain Privileged User Account (PUA) certification
- Experience with Kubernetes
- Experience with GitLab pipelines
- Experience with Linux
- Experience with containerized environments
- Experience supporting enterprise scale production systems
- Experience with cloud services (preferably AWS) and cloud infrastructure
- Familiarity with Elasticsearch
- Familiarity with PostgreSQL
- Familiarity with Logstash
- Familiarity with Kibana
- Familiarity with Keycloak
- Demonstrated success in cross functional coordination and execution
- Strong communication skills
- The ability to perform under pressure during incidents
Desired Qualifications
- Experience with Agile methodologies
- Experience with creating customized dashboards to track SLIs and other key performance indicators
- Development experience (Bash, PowerShell, SALT, Python, Groovy, Java, etc.)
- Experience with Appian or other low‐code platforms
- Experience with technologies such as Kafka
- Experience with technologies such as AMQP/JMS
- Experience with technologies such as Prometheus/Grafana
- Experience with GPU‐based Kubernetes
- Experience with SALT automation
- Experience with Nexus
- Experience with GraphQL
- Knowledge of security best practices (authN/Z, secrets management, data protection)
- Infrastructure‐as‐code experience (CloudFormation, Terraform, Pulumi)
- AWS cloud certifications
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.