Deep Learning Compiler CI/Infrastructure Engineer
On-siteShanghai, Shanghai, China
Job Summary
Deep Learning Compiler CI/Infrastructure Engineer at NVIDIA responsible for designing and operating scalable CI/CD infrastructure powering NVIDIA's deep learning compiler stacks across GPU and accelerator environments. Build, maintain, and improve CI pipelines to support development, verification, and release of deep learning compiler stacks; enhance CI reliability and signal quality; apply automation and AI/agent-based workflows to reduce manual CI operations; build reusable, self-service CI platforms that support multiple products, projects, model suites, hardware targets, and software configurations in collaboration with compiler, infrastructure, and release teams.
Required Qualifications
- BS, MS, or PhD in Computer Science, Computer/Electrical Engineering, Mathematics, or related field
- 5+ years of experience designing, scaling, and operating CI/CD, build/release, or developer infrastructure for complex software systems
- Proven experience building CI platforms end-to-end using GitLab CI, Jenkins, or similar tools, including pipeline orchestration, compute/runner management, artifact and package systems, and observability
- Strong software engineering skills (Python required)
- Familiarity with edge devices (SOC, e.g. NVIDIA Tegra) in host-target architecture and automation nuances
- Proven track record of designing, building, and deploying AI/LLM-based systems in real engineering workflows
- Experience with multi-GPU / multi-node workloads (Slurm, Kubernetes, cloud)
- Knowledge of compiler IRs and infrastructure (LLVM/MLIR, TensorRT IR) for testing and debugging
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.