Role at a glance
- Salary
- $148K – $276K/yr
- Location
- Santa Clara, California, United States
- Work arrangement
- On-site
- Employment
- Full-time
- Experience
- 5+ years of overall experience in DevOps, SRE, or Systems Integration roles.
- Education
- Bachelor Science Degree in Computer Science or similar academic degree, or equivalent experience.
Spotted an issue?
We’ll check it against the original posting.
Role Summary
This role solves software integration challenges for NVIDIA’s next-generation data center platforms, supporting GPU architectures and AI infrastructure involving Ethernet, InfiniBand, virtualization, GPUs, network stacks, firmware, and drivers. The position provides first-tier technical support to R&D teams and helps stabilize software deployments for high-speed communication and system management projects.
What You'll Do
- Fix and prioritize complex systems during high-stakes bringups and Proofs of Concepts for next-generation computing architectures.
- Manage the integration of large-scale products involving GPUs, network stacks, firmware, and drivers.
- Create, recreate, and redeploy software artifacts, including fixing code, updating builds, and providing workarounds.
- Serve as the primary technical point of contact for R&D teams resolving infrastructure and integration blockers.
- Work with R&D, Verification, and DevOps teams to streamline CI/CD pipelines for high-speed interconnect and system management projects.
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Proven software engineering background with a deep understanding of standard methodologies in software development, modern Linux-based operating systems, and computer networking. Deep knowledge of Linux distributions (Ubuntu/RHEL), Docker containerization, and coding in C/C++, Python, and Bash.
Required
- Software engineering background
- Software development methodologies
- Linux-based operating systems
- Computer networking
- Linux distributions (Ubuntu/RHEL)
- Docker containerization
- C/C++
- Python
Preferred
- High-performance networking (InfiniBand, Ethernet)
- gRPC
- gNMI
- REST
- JSON
- Large-scale HW+SW converged systems, such as rack-scale computing or GPU clusters
Original job description
Content provided by the employer
Original job description
Content provided by the employer
NVIDIA is looking for an outstanding candidate to solve SW integration challenges for our next-generation data center platforms. You will be at the heart of our latest GPU architectures and advanced AI infrastructure projects, ensuring the seamless integration of world-class technologies in the areas of High-Speed Communication and virtualization. You will support products that leverage Ethernet and InfiniBand protocols, delivering a broad range of advanced compute and networking technologies for the world's most demanding AI workloads. In this role, you will provide first-tier support to R&D teams, acting as the bridge between pioneering hardware and stable software deployments.
What you’ll be doing:
Fixing and prioritising complex systems during high-stakes bringups and Proof of Concepts (PoCs) for next-generation computing architectures.
Managing the integration of large-scale products involving GPUs, complex Network Stacks, Firmware, and Drivers.
Creating, recreating, and redeploying software artifacts. You will be responsible for fixing code, updating builds, or providing creative workarounds to unblock development.
Serving as the primary technical point of contact for R&D teams to resolve immediate infrastructure and integration blockers.
Working closely with R&D, Verification, and DevOps teams to streamline the CI/CD pipeline for specialized high-speed interconnect and system management projects.
What we need to see:
Bachelor Science Degree in Computer Science or similar academic degree, or equivalent experience.
Proven software engineering background with a deep understanding of standard methodologies in software development, modern Linux-based operating systems, and computer networking.
5+ years of overall experience in DevOps, SRE, or Systems Integration roles.
Deep knowledge of Linux distributions (Ubuntu/RHEL) and containerization using Docker.
Coding skills in C/C++, Python, and Bash for automation and system-level fixes.
Experience with GitLab and GitLab CI for managing complex build pipelines.
Ability to multi-task, self-manage in a fast-paced environment, and lead technically during critical system failures.
Excellent problem-solving and critical thinking abilities.
Ways to stand out from the crowd:
In depth knowledge and familiarity with high-performance networking (InfiniBand, Ethernet).
Practical experience with gRPC, gNMI, REST, and JSON for system management and telemetry.
A proven track record of working on large-scale HW+SW converged systems (e.g., rack-scale computing or GPU clusters).
With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 148,000 USD - 235,750 USD for Level 3, and 176,000 USD - 276,000 USD for Level 4.You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About the company
NVIDIA
Large Enterprise
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.