Senior Product Manager - AI Platform Inference
On-siteSanta Clara, California, United States
Job Summary
Create products to help developers build better Inference deployments on NVIDIA GPUs. Develop product strategy, roadmaps, and go-to-market plans while collaborating with internal and external developers to build product-based roadmaps for model optimization software. Work with leadership to align with and drive company strategy, ensuring tools, SDKs, and libraries enable successful AI deployments. This role requires 6+ years of technical product management experience and knowledge of GenAI, performance optimization, and software delivery. The position supports NVIDIA's goal of enabling deep learning across all GPU use cases.
Required Qualifications
- Experience with Inference deployment and optimization software (ex. vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO, etc.)
- Demonstrable knowledge of GenAI or machine learning concepts, particularly around performance optimization, and software development and delivery
- BS or MS degree in Computer Science, Computer Engineering, or similar experience (or equivalent experience)
- 6+ years of technical product management, or similar, experience at a technology company
- Strong communication and interpersonal skills
Desired Qualifications
- Experience leading optimization products for Inference
- Working on Open Source & Github-first developer products with deep customer interactions
- Knowledge of GPU architecture, HW/SW co-design, and performance profiling
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.