EDA Computing / HPC Infrastructure MTS
$104,000–$175,000 year
On-siteAustin, Texas, United States or Essex Junction, Vermont, United States
Job Summary
Provide senior-level support for EDA compute clusters handling simulation, physical verification, OPC, lithography, signoff, and tapeout workloads. Operate, optimize, and troubleshoot IBM LSF batch scheduling environments while analyzing job failures, queue bottlenecks, license constraints, and node issues. Deliver technical ownership for hybrid HPC platforms across on-premises infrastructure and AWS, managing storage services like NFS, NetApp ONTAP, AWS FSx, and S3. Design automation tools to enhance operations, observability, and support efficiency, while owning complex ServiceNow incidents and leading root-cause analysis. Contribute to team meetings, knowledge sharing, and continuous improvement activities. Minimum 8+ years of experience in EDA computing, HPC, Linux infrastructure, or production engineering support required.
Required Qualifications
- Bachelor's degree in Computer Science, Engineering, Information Technology, or equivalent practical experience
- Minimum 8+ years of relevant experience for MTS, or 10+ years of relevant experience for Senior MTS, in EDA computing, HPC, Linux infrastructure, storage, cloud, or production engineering support
- Strong experience supporting EDA computing, HPC, or large-scale Linux production environments
- hands-on experience with batch schedulers, preferably IBM LSF
- strong Linux system administration skills on RHEL-based platforms
- Bash scripting capability
- ability to analyze complex incidents, identify root causes, and implement sustainable technical solutions
- Fluency in English Language – written & verbal
Desired Qualifications
- Master's degree in Computer Science, Engineering, Information Technology, or related technical field preferred
- Experience supporting semiconductor design and tapeout flows, including Calibre, ASML Brion, Cadence, Synopsys, or similar EDA applications
- AWS or other cloud platform experience for HPC or engineering workloads
- NetApp ONTAP, NFS, FSx, S3, or high-performance file storage experience
- Python automation experience
- experience with monitoring, logging, dashboarding, capacity reporting, and operational analytics tools
- Project management skills – i.e., the ability to innovate and execute on solutions that matter
- the ability to navigate ambiguity
- Strong written and verbal communication skills
- Strong planning & organizational skills
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.