NVIDIA

NVIDIA

Posted via Workday

Senior Deep Learning Algorithm Engineer

Posted Aug 6, 2026

Role at a glance

Salary
$152K – $287.5K/yr
Location
2 Locations, California, United States
Work arrangement
On-site
Employment
Full-time
Experience
3+ years building, profiling, and debugging performance-critical distributed or ML systems.
Education
BS, MS, PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related field (or equivalent experience).

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Senior Deep Learning Algorithms Engineer will advance Dynamo, NVIDIA’s open-source distributed inference platform for large-scale, low-latency AI services. The role works across research, software, systems, and hardware teams and engages with open-source frameworks and external partners to improve AI inference performance, efficiency, and deployment.

What You'll Do

  • Design, build, and maintain Dynamo integrations for open source frameworks vLLM, SGLang, TRTLLM.
  • Partner with open source communities to land measurable gains in latency, throughput, reliability, and efficiency.
  • Showcase NVIDIA token/watt leadership by pushing the pareto frontier on public/private benchmarks
  • Find and remove bottlenecks across runtimes, kernels, networking, routing, and orchestration.
  • Develop inference optimizations for scheduling, disaggregation, KV caching, and autoscaling.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Strong programming skills in Python and/or Rust, C++. Understanding of modern ML architectures and inference techniques.

Required

  • Strong programming skills in Python and/or Rust, C++
  • Understanding of modern ML architectures and inference techniques
  • 3+ years building, profiling, and debugging performance-critical distributed or ML systems
  • BS, MS, PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related field (or equivalent experience)

Preferred

  • High agency and a track record of leading ambiguous work end to end
  • Experience with AI Accelerators
  • Open-source contributions / leadership
  • Research in ML inference or distributed systems

Original job description

Content provided by the employer

NVIDIA is seeking a Senior Deep Learning Algorithms Engineer to advance Dynamo, our open-source distributed inference platform for large-scale, low-latency AI services. You’ll lead architecture and performance work across Dynamo and open source frameworks. You’ll collaborate across research, software, systems, and hardware teams to make AI inference faster, more efficient, and easier to deploy. You’ll engage with the broader ecosystem, including vLLM, SGLang, and TensorRT-LLM as well as with external partners to build the best operating system for AI. If you’re excited by deep learning, performance engineering, and distributed systems, we’d love to hear from you.
 

What you'll be doing:

  • Design, build, and maintain Dynamo integrations for open source frameworks vLLM, SGLang, TRTLLM.

  • Partner with open source communities to land measurable gains in latency, throughput, reliability, and efficiency.

  • Showcase NVIDIA token/watt leadership by pushing the pareto frontier on public/private benchmarks

  • Find and remove bottlenecks across runtimes, kernels, networking, routing, and orchestration.

  • Develop inference optimizations for scheduling, disaggregation, KV caching, and autoscaling.
     

What we need to see:

  • BS, MS, PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related field (or equivalent experience).

  • 3+ years building, profiling, and debugging performance-critical distributed or ML systems.

  • Strong programming skills in Python and/or Rust, C++.

  • Understanding of modern ML architectures and inference techniques
     

Ways to stand out from the crowd:

  • High agency and a track record of leading ambiguous work end to end.

  • Experience with AI Accelerators 

  • Open-source contributions / leadership

  • Research in ML inference or distributed systems.
     

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

#LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 9, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA

About the company

NVIDIA

Large Enterprise

NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.