Santa Clara University logo
Santa Clara UniversityPosted 1 month ago

Research Computing Engineer

$115,200–$129,600 year

HybridSanta Clara, California, United States

Full TimeLarge

Job Summary

Lead the strategic roadmap and long-term capacity planning for Santa Clara University's High-Performance Computing (HPC) infrastructure, partnering with academic stakeholders to forecast computational demands. Own the full-cycle consultation process with researchers and faculty, translating complex academic requirements into scalable technical solutions while architecting proactive infrastructure enhancements for AI, Machine Learning, and GPU-accelerated processing. Design standard operating procedures, automated workflows, and lifecycle management processes for scientific software deployment, alongside comprehensive training programs in modern code-management and version control. Manage the deployment lifecycle, configuration, and optimization of specialized scientific software, containerized environments, and shared libraries, including workload managers and cluster schedulers. Enforce comprehensive security frameworks, monitor system telemetry for root-cause analysis, and maintain high-speed network fabrics within the HPC environment.

Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, or a highly quantitative field
  • 8–10 years of progressively responsible experience in Information Technology operations and system design
  • Experience explicitly leading, architecting, and supporting multi-node HPC cluster environments
  • Advanced, hands-on mastery of Linux systems administration
  • Demonstrated experience writing and debugging complex scripts in Bash, Python, or Ansible
  • Deep knowledge of workload managers (Slurm)
  • Deep knowledge of container technologies (Docker, Apptainer)
  • Proven success implementing distributed file systems (BeeGFS, Lustre)
  • Advanced understanding of cybersecurity principles, encryption standards, and risk-mitigation strategies unique to open research cluster environments
  • Exceptional interpersonal and verbal communication skills
  • Strong analytical skills with a proactive approach to identifying and resolving technical and human issues
  • Ability to independently design, implement, and govern enterprise-grade computational environments and workflows
  • Ability to meet in-person with researchers and colleagues on the Santa Clara University campus
  • Regular on-site presence required, typically at least 3–4 days a week
  • Occasional evening or weekend work required for system maintenance or outage mitigation
  • Ability to lift or move objects up to 50 pounds

Desired Qualifications

  • Advanced degree (MS or PhD)
  • 5+ years of experience explicitly leading, architecting, and supporting multi-node HPC cluster environments

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce