Trener Robotics logo
Trener RoboticsPosted 1 week ago

Senior Cloud Infrastructure Engineer

On-siteSan Jose, California, United States

Full TimeSenior LevelSmall

Job Summary

Own cloud foundations across Google Cloud and AWS, including project topology, landing zones, organization policies, and baseline guardrails. Manage cross-cloud networking, VPCs, routing, peering, DNS, and private connectivity while authoring reusable Terraform/OpenTofu modules for infrastructure lifecycle. Implement cloud IAM, workload identity, and access controls in partnership with Security, ensuring SOC 2 compliance through resource labeling and configuration standards. Define Kubernetes cluster lifecycle, baseline configurations, and shared services beneath application workloads. Establish infrastructure cost visibility via tagging, labeling, budgets, and alerts to surface waste. Operate as a senior individual contributor without people management, focusing on establishing a deliberate multi-cloud foundation for Trener's industrial robotics platform.

Required Qualifications

  • Proven experience operating production infrastructure as a cloud, infrastructure, platform, or site reliability engineer
  • Terraform or OpenTofu at module-authoring depth - writing and versioning reusable modules, managing state across environments, and handling the lifecycle of real infrastructure over time
  • Cloud foundation design experience - multi-account, multi-project, landing-zone, guardrail, IAM, or network architecture beyond a handful of isolated workloads
  • Experience with infrastructure composition or orchestration - Atmos, Terragrunt, Terraspace, or an equivalent approach to managing reusable infrastructure across environments
  • Strong networking depth - VPCs, routing, peering, DNS, TLS, load balancing, firewall policy, IP addressing, and connectivity troubleshooting
  • Kubernetes working knowledge - cluster operations, RBAC, networking, shared services, and troubleshooting workloads that will not start or communicate correctly
  • Helm chart authoring, not just chart installation
  • Production cloud experience across at least two major providers, or deep experience with one plus substantial hands- on involvement extending an organization into another
  • Python and Bash for automation, CLIs, diagnostics, and glue
  • Linux fluency
  • Experience with controlled infrastructure environments - SOC 2, HIPAA, HITRUST, FedRAMP, PCI, or similar security/compliance expectations
  • Fluency in English and strong written communication
  • Architectural decisions here are expected to be documented and reviewed in writing

Desired Qualifications

  • GitOps workflows across multiple environments and comfort operating infrastructure through reviewed, auditable changes
  • Externalized secrets management and Kubernetes-native secret delivery patterns
  • Operating Prometheus/Grafana/Loki or comparable observability tooling as an infrastructure consumer and operator
  • Experience introducing or formalizing a second cloud, including a clear account of what you would repeat and what you would change
  • Cloud identity federation and workload identity across providers
  • Disaster-recovery, regional resilience, or high-availability infrastructure work
  • Private cloud networking or overlay connectivity

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce