NVIDIA

NVIDIA

Posted via Workday

Senior Research Engineer - Enterprise Products

Posted Aug 5, 2026

Role at a glance

Salary
$192K – $356.5K/yr
Location
3 Locations, Washington, United States
Work arrangement
On-site
Employment
Full-time
Experience
8+ years of industry experience in Deep Learning frameworks (PyTorch or TensorFlow).
Education
Bachelor's of Master's degree in Computer Science or equivalent experience.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Senior Research Engineer will develop optimized generative AI inference technologies for NVIDIA’s accelerated serving stack. The role spans applied research, engineering, evaluation, open-source development, and collaboration with research and engineering teams to improve how generative AI systems are integrated and deployed.

What You'll Do

  • Design and evaluate routing policies for LLM traffic across mixture-of-model systems.
  • Build and run agentic benchmarks to measure algorithm quality and create calibration data and routing profiles.
  • Ship design documents, code reviews, documentation, and community contributions to an open-source repository.
  • Collaborate with engineering teams across NVIDIA to integrate software with the NVIDIA accelerated serving stack.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Computer Science bachelor's or master's degree or equivalent experience; 8+ years of industry experience with PyTorch or TensorFlow; experience designing or running LLM evaluations or benchmarks; understanding of machine learning, deep neural networks, natural language processing, or speech recognition; empirical research experience; computer science fundamentals in algorithms, data structures, computational complexity, parallel and distributed computing, and system software; strong communication and interpersonal skills.

Required

  • Computer Science bachelor's or master's degree or equivalent experience
  • 8+ years of industry experience in PyTorch or TensorFlow
  • LLM evaluation or benchmark experience
  • Understanding of machine learning, deep neural networks, natural language processing, or speech recognition
  • Empirical research experience
  • Algorithms and data structures
  • Computational complexity
  • Parallel and distributed computing

Preferred

  • Experience architecting or developing large-scale distributed systems for deep learning
  • Agentic benchmark creation and publications
  • Knowledge of CPU and/or GPU architecture
  • GPU programming (CUDA)
  • History of mentoring junior engineers and interns

Original job description

Content provided by the employer

We are now looking for a Senior Research Engineer passionate about Generative AI inference. Are you excited to change the way people infuse AI into products and services? NVIDIA is at the forefront of generative AI models, from language to images. NVIDIA provides building blocks to democratize AI and make generative AI easy to develop, integrate, and deploy. Our team is dedicated to developing optimized inferencing technologies to support our growing generative AI needs. We contribute to all steps of the machine learning lifecycle: from conceptualization, to applied research, engineering for optimized inference, and deployment. Collaborate with research teams, engineers, and open-source community.

What you will be doing:

  • Design and evaluate routing policies for LLM traffic to best use mixture of model systems.

  • Build and run agentic benchmarks (e.g., Terminal-Bench ) to measure algorithm quality, and turn results into calibration data and routing profiles

  • Ship to an open-source repo: design docs, code review, docs, and community contributions

  • Collaborating with engineering teams across all of NVIDIA to ensure our software integrates seamlessly up and down the NVIDIA accelerated serving stack.

What we need to see:

  • Bachelor's of Master's degree in Computer Science or equivalent experience.

  • 8+ years of industry experience in Deep Learning frameworks (PyTorch or TensorFlow).

  • Experience designing or running LLM evaluations/benchmarks — ideally agentic ones — and drawing statistically sound conclusions from them

  • Understanding of modern techniques in Machine Learning, Deep Neural Networks, Natural Language Processing, or Speech Recognition.

  • Empirical research mindset: forming hypotheses about new algorithms, running calibrations, iterating on results

  • Strong communication and interpersonal skills, along with the ability to work in a dynamic and distributed team. A history of mentoring junior engineers and interns is a huge plus.

  • A desire to constantly grow and learn new things.

  • Strong computer science fundamentals - algorithms and data structures, computational complexity, parallel and distributed computing, system software.

Ways to stand out from a crowd:

  • Experience architecting or developing large-scale distributed systems for deep learning.

  • Agentic benchmark creation and publications.

  • Knowledge of CPU and/or GPU architecture.

  • GPU programming (CUDA).

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 192,000 USD - 304,750 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 14, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA

About the company

NVIDIA

Large Enterprise

NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.