Nscale logo
NscalePosted 3 weeks ago

Director, HPC Systems Software Engineering

$230,000–$343,333 year

On-siteHouston, Texas, United States

Part TimeSenior LevelSmall

Job Summary

Define and evolve strategy for a major HPC systems engineering area aligned with business goals, customer needs, and long-term service direction. Translate company strategy into clear priorities, investment areas, and execution plans while ensuring architecture, operational maturity, and delivery remain aligned across teams. Lead a major HPC systems engineering organization through managers and senior technical leaders, shaping team topology, ownership boundaries, and operating mechanisms to support effective execution at scale. Own delivery outcomes, execution quality, and service quality, removing systemic blockers that reduce engineering effectiveness or create avoidable delivery risk. Hire, develop, and retain strong engineering leaders and senior engineers, raising the leadership bar through coaching, feedback, and succession planning. Serve as a trusted partner to senior engineering, product, infrastructure, security, and business stakeholders, representing the organization in strategic planning, executive discussions, and partner engagements. Own budget, headcount planning, and organizational investment decisions to position the HPC systems engineering function to support broader company goals.

Required Qualifications

  • Proven experience leading a significant HPC systems, platform engineering, or infrastructure-heavy engineering organisation
  • Strong technical credibility in software engineering for HPC systems, distributed infrastructure, Linux-based platforms, and production operations
  • Demonstrated ability to operate at org scope, translating business and engineering strategy into team structure, priorities, and execution
  • Strong track record of leading through managers and senior technical leaders rather than relying primarily on individual contribution
  • Experience building high-performing engineering organisations, including hiring, leadership development, organisation design, and performance management
  • Strong judgement in balancing execution, platform quality, operational risk, technical quality, and long-term investment
  • Experience working closely with senior ICs to turn HPC platform strategy into practical delivery outcomes
  • Strong communication and stakeholder management skills, including the ability to explain priorities, trade-offs, and constraints clearly to senior leadership
  • Experience owning budget, headcount, and planning decisions for a meaningful engineering area
  • Strong understanding of GPU infrastructure, HPC scheduler environments, networking constraints, and production service maturity

Desired Qualifications

  • Experience with Slurm
  • Experience with Kueue or other batch scheduling technologies

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce