Role at a glance
- Salary
- $168K – $327.8K/yr
- Location
- Santa Clara, California, United States
- Work arrangement
- On-site
- Employment
- Full-time
- Experience
- 6+ years of technical product management, or similar, experience at a technology company
- Education
- BS or MS degree in Computer Science, Computer Engineering, or similar experience (or equivalent experience)
Spotted an issue?
We’ll check it against the original posting.
Role Summary
The Senior Product Manager for AI Platform Inference will build tools, SDKs, and libraries that enable developers to deploy inference workloads on NVIDIA GPUs. The role works with internal and external developers, NVIDIA leadership, and marketing teams to improve model optimization software and define product strategy and go-to-market plans.
What You'll Do
- Create products to help developers build better Inference deployments
- Develop product strategy, roadmaps, and go-to-market plans
- Collaborate with internal and external developers to build product-based roadmaps for model optimization software
- Work with leadership to align with and drive company strategy
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Experience with inference deployment and optimization software; demonstrable knowledge of GenAI or machine learning concepts, particularly around performance optimization, and software development and delivery; strong communication and interpersonal skills.
Required
- Experience with Inference deployment and optimization software (ex. vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO, etc.)
- Demonstrable knowledge of GenAI or machine learning concepts, particularly around performance optimization, and software development and...
- BS or MS degree in Computer Science, Computer Engineering, or similar experience (or equivalent experience)
- 6+ years of technical product management, or similar, experience at a technology company
- Strong communication and interpersonal skills
Preferred
- Experience leading optimization products for Inference
- Working on Open Source & Github-first developer products with deep customer interactions
- Knowledge of GPU architecture, HW/SW co-design, and performance profiling
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Inference is the fastest growing and most competitive area in Generative AI today. It is where AI models impact our daily life, and where ever bit of accuracy and performance matters for quality, safety, and cost. Inference is also constantly evolving, with new acceleration algorithms, usecases, and deployment techniques. As a Senior Product Manager for AI Platform Inference you will be responsible for building the tools, SDKs, and libraries which enables developers' Inference deployments to thrive on NVIDIA GPUs.
As NVIDIA Product Managers, our goal is to enable developers to be successful on the NVIDIA Platform, and push the boundaries of what is possible with their AI deployments! For Inference, we are the champions inside NVIDIA for AI developers looking to accelerate their deployments on GPUs. We work directly with developers inside and outside of the company to identify key improvements, create roadmaps, and stay alert on the inference landscape. We also work with NVIDIA leaders to define clear product strategy, and marketing team teams to build go-to-market plans. The Product Management organization at NVIDIA is a small, strong, and impactful group. We focus on enabling deep learning across all GPU use cases and providing great solutions for developers. We are seeking a rare blend of product skills, technical depth, and passion to make NVIDIA great for developers. Does that sounds familiar? If so, we would love to hear from you!
What you'll be doing:
Create products to help developers build better Inference deployments
Develop product strategy, roadmaps, and go-to-market plans
Collaborate with internal and external developers to build product-based roadmaps for model optimization software
Work with leadership to align with and drive company strategy
What we need to see:
Experience with Inference deployment and optimization software (ex. vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO, etc.)
Demonstrable knowledge of GenAI or machine learning concepts, particularly around performance optimization, and software development and delivery
BS or MS degree in Computer Science, Computer Engineering, or similar experience (or equivalent experience)
6+ years of technical product management, or similar, experience at a technology company
Strong communication and interpersonal skills
Ways to stand out from the crowd:
Experience leading optimization products for Inference
Working on Open Source & Github-first developer products with deep customer interactions
Knowledge of GPU architecture, HW/SW co-design, and performance profiling
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About the company
NVIDIA
Large Enterprise
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.