Binance logo
BinancePosted 1 month ago

Senior Infrastructure Engineer — IDC / Bare-Metal Kubernetes

On-siteTokyo, Tokyo, Japan or Hong Kong, Hong Kong

Full TimeSenior LevelLarge

Job Summary

Design rack layout, network topology, and overall configuration for colocation while implementing an isolated Out-of-Band (OOB) management network with full security hardening. Coordinate with colocation and hardware vendors for servers, switches, and remote hands, then build bare-metal automation for fleet-wide zero-touch OS provisioning. Set up production-grade self-hosted Kubernetes clusters from scratch, maintain the core stack including Cilium and GitOps, and own cluster lifecycle management for upgrades and incident response. Establish hybrid-cloud interconnect between IDC and public cloud, paving the way for scale-out GPU and VM workloads. Participate in 24×7 on-call rotation and write SOPs, runbooks, and post-mortems.

Required Qualifications

  • 5+ years in data center / infrastructure / platform engineering
  • Hands-on experience building physical infrastructure from scratch: rack layout, network topology, server commissioning, and coordinating cross-connects and remote hands with colo / vendors
  • Practical experience designing and operating OOB management networks (BMC, IPMI, Redfish)
  • Have stood up production-grade self-hosted Kubernetes from scratch, and can independently debug cluster-level issues (CNI, CSI, storage)
  • Strong Linux systems administration and performance tuning (kernel, networking, storage I/O)
  • Bare-metal automation experience with at least one of: MAAS, Tinkerbell, Cluster API
  • Proficient with Terraform, Ansible, and at least one scripting language (Python / Go / Bash)
  • Experience with Cisco network and related techniques (VLAN, LACP/LAG, BGP, ACL, etc)
  • Experience with Palo Alto firewall configuration
  • Experience with storage systems (NetApp, Dell EMC, Pure Storage)
  • Fluent in English or Mandrain

Desired Qualifications

  • High-density racks (30kW+) and 400G+ networking experience
  • Familiarity with immutable OS (Talos Linux / Flatcar / Bottlerocket)
  • Proficiency across both AWS and GCP; cross-cloud data migration experience
  • Experience building storage or Bigdata / offline data clusters
  • Virtualization experience (KubeVirt or similars)
  • Exposure to NVIDIA GPU Operator and K8s GPU workloads
  • CKA / CKS certification
  • CCNP / CCIE certification
  • Professional-level Japanese

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce