Senior Performance Engineer
RemoteZürich, Zurich, Switzerland or Switzerland
Job Summary
Profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clusters. Explore performance characteristics of high-performance networking and collective communications, including NCCL, RDMA, and MPI. Identify bottlenecks across networking, compute, memory, and system architecture while developing diagnostic tools and defining test plans for new technologies. Collaborate with hardware, firmware, and software teams to collect and refine telemetry data, ensuring high standards for data quality and reproducibility. This role sits within NVIDIA's fast-paced R&D environment, supporting supercomputing and AI infrastructure that powers global intelligence.
Required Qualifications
- B.Sc. or M.Sc. in Computer Science, Computer Engineering, Software Engineering, or equivalent experience
- 5+ years of experience in performance analysis, systems engineering, or HPC/AI infrastructure
- Demonstrated expertise in performance analysis skills and methodologies
- Hands-on experience with high-performance networking (RDMA, MPI, NCCL, congestion control)
- Strong understanding of system performance metrics (latency, throughput, resource utilization)
- Exposure to hardware, firmware, or embedded telemetry environments
- Strong analytical, problem-solving, and communication skills
- Ability to work effectively in cross-functional, fast-paced R&D teams
Desired Qualifications
- Knowledge of CUDA, NCCL internals, and congestion control algorithms
- Deep system-level understanding of CPU architectures, GPUs, HCAs, memory, and PCIe
- Experience with NVIDIA GPUs, CUDA, and deep learning frameworks such as PyTorch or TensorFlow
- Experience with cloud platforms
- Proficiency in Python
- experience with Bash and C/C++
- strong experience working in Linux environments
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.