Lambda logo
LambdaPosted 1 month ago

Senior HPC Platform Hardware Engineer

$255,000–$340,000 year

RemoteUnited States

Full TimeSenior LevelSmall

Job Summary

Lead integration of OEM and white-label HPC AI/ML, compute, storage, and network hardware into Lambda's reference architectures. Drive end-to-end new product introduction, including system bring-up, vendor engagement, production readiness, and risk closure. Identify and resolve hardware issues across domains during NPI, partnering with architects to translate blueprints into concrete configurations. Collaborate with supply chain, quality, and fleet teams to evaluate vendors, review BOMs, and ensure on-time delivery of gigawatt-scale AI infrastructure. Serve as technical lead for lab prototyping and enablement of new hardware systems. This role requires presence in the San Jose office four days per week.

Required Qualifications

  • 5 years of technical lead experience on hardware NPI and deployment for HPC, data center, or cloud infrastructure products
  • familiar with hardware NPI processes
  • deep knowledge and hands-on experiences in one or many of the following hardware platforms: AI/ML, general compute (x86 and ARM), storage systems, or network switches
  • Broad hardware engineering domain knowledge in one or many of the below areas: electrical, thermal, mechanical, power, signal integrity, safety, compliance, reliability and manufacturing
  • Are comfortable working hands-on in labs to enable and bring up new hardware systems
  • Experiences in identifying, triaging and root causing hardware issues during NPI and at scale in the fleet
  • Experience in PLM systems and BOM structure
  • Collaborate well cross functionally to deliver production-ready hardware solutions
  • Strong ownership and can do attitude, self-starter who feels comfortable working in ambiguity
  • presence in our San Jose office location 4 days per week

Desired Qualifications

  • 10+ years of technical lead experience on hardware NPI and deployment for HPC, data center, or cloud infrastructure products
  • Experience supporting AI/ML infrastructure and accelerated compute hardware (e.g., NVIDIA, AMD, Intel)
  • Experience in rack scale server development and liquid cooling designs
  • Exposure to fleet observability, BMC/BIOS/Network configuration and automation
  • Background in performance tuning, benchmarking, and systems validation workflows
  • Can interpret platform-level architecture requirements and select or adapt OEM and white-label solutions to fit
  • Prior experience contributing to reference designs or large-scale infrastructure blueprints
  • Are experienced with vendor-led product development cycles and can drive hardware evaluation, risk mitigation, and feedback into roadmap decisions

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce