Senior Hardware Systems Engineer
$170,000–$205,000 year
On-siteSunnyvale, California, United States
Job Summary
Drive the full hardware development and sustaining lifecycle from prototype bring-up to large-scale production, owning debugging, validation, and reliability for GPU- and CPU-based infrastructure. Develop automation frameworks for hardware testing and diagnostics while leading deep troubleshooting across PCIe, InfiniBand, and NVMe/storage systems. Conduct rigorous system characterization for high-performance compute platforms and collaborate with mechanical, thermal, firmware, and software teams to resolve system-level issues. Support E2E integration testing to ensure products meet scalability expectations and drive prototyping for high-volume manufacturing. Provide data-driven insights to influence the hardware roadmap and identify opportunities for new technologies and sustainability improvements aligned with long-term objectives.
Required Qualifications
- 8–10+ years of experience in hardware development, validation, sustaining engineering, or production engineering
- Strong hands-on expertise in PCIe, InfiniBand, and NVMe/storage debugging and development
- Deep proficiency in hardware bring-up, board-level debugging, and system-level validation
- Ability to design and implement automation frameworks for hardware testing (Python, Shell, or similar)
- Technical background in digital and analog design, server architecture, and high-performance compute hardware
- Experience working across thermal, mechanical, firmware, and software functions in multidisciplinary environments
- Strong analytical and problem-solving skills with a data-driven approach
- Excellent communication and collaboration skills for working with internal teams and external partners
- Bachelor's or Master's degree in Electrical Engineering, Computer Engineering, or equivalent experience
Desired Qualifications
- Experience designing or optimizing GPU-to-GPU communication architectures for AI/ML workloads
- Direct experience integrating NVLink or other next-generation GPU interconnect technologies
- Familiarity with cutting-edge GPU architectures and how to leverage them in AI/HPC environments
- Expertise supporting or designing systems across both ARM and x86 server architectures
- Background in sustainable or energy-efficient hardware design practices
- Advanced certifications or coursework in AI/HPC hardware systems
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.