NVIDIA

NVIDIA

Posted via Workday

Senior Developer Technology Engineer - Edge Agentic AI

Posted Aug 5, 2026

Role at a glance

Salary
$152K – $287.5K/yr
Location
Santa Clara, California, United States
Work arrangement
On-site
Employment
Full-time
Experience
5+ years of professional experience in local GPU deployment, profiling and optimization.
Education
A Bachelor's or Master's degree or equivalent experience in Computer Science, Engineering, or a related field.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Developer Technology Engineer works with internal teams, external app developers, and enterprise ISVs to enable agentic AI workflows at the edge on NVIDIA RTX and DGX platforms. The role focuses on deployment performance, open-source GenAI software improvements, and collaboration that informs future GPU features.

What You'll Do

  • Solve local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX and DGX platforms with internal and external partners
  • Use profiling and debugging tools to analyze accelerated agentic AI workflows and identify insufficient system utilization
  • Conduct hands-on trainings, develop sample code, and host presentations on efficient agentic AI deployment
  • Improve features and performance of open-source software including GGML, Llama.cpp, Ollama, vLLM, and ONNX Runtime
  • Collaborate with GPU driver, architecture, and research teams to influence next-generation GPU features
  • Provide technical leadership and mentorship to junior engineers

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Requires professional experience in local GPU deployment, profiling, and optimization; programming proficiency in C/C++ and Python; Windows and Linux development experience; and experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.

Required

  • 5+ years of professional experience in local GPU deployment, profiling and optimization
  • Strong proficiency in C/C++ and Python
  • Familiarity with and development experience on Windows and Linux
  • Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite
  • Strong problem-solving skills
  • Excellent interpersonal and communication skills

Preferred

  • Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, NVIDIA Model Optimizer
  • Expertise with professional agentic AI use cases, i.e., digital content creation and productivity workflows
  • Experience working with open-source LLM and GenAI software
  • Detailed knowledge of the latest generation GPU architectures
  • Experience with AI deployment on NPUs and ARM architectures

Original job description

Content provided by the employer

At NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

As a Developer Technology Engineer, you will be at the forefront of innovation, working with leading industry partners and pioneering open-source projects to enable professional agentic AI workflows at the edge powered by NVIDIAs RTX and DGX platforms. This role offers an outstanding opportunity to collaborate with world-class talent and make a significant contribution to the evolving landscape of enterprise and consumer agentic AI.

What you'll be doing:

  • Work closely across internal engineering and product teams as well as external app developers and enterprise ISVs on solving local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX & DGX.

  • Apply powerful profiling and debugging tools for analyzing most demanding accelerated end-to-end agentic AI workflows to detect insufficient system utilization resulting in suboptimal runtime performance.

  • Conduct hands-on trainings, develop sample code and host presentations to give good guidance on efficient end-to-end agentic AI deployment targeting optimal runtime performance.

  • Improve LLM & GenAI user experience by working on feature and performance enhancements of OSS software, including but not limited to projects like GGML, Llama.cpp, Ollama, vLLM, ONNX Runtime.

  • Collaborate with GPU driver and architecture teams as well as NVIDIA research to influence next generation GPU features by providing real-world workflows and giving feedback on partner and customer needs.

  • Providing technical leadership and mentorship to junior engineers, encouraging an inclusive and high-performing team environment.

What we need to see:

  • A proven track record 5+ years of professional experience in local GPU deployment, profiling and optimization.

  • A Bachelor's or Master's degree or equivalent experience in Computer Science, Engineering, or a related field.

  • Strong proficiency in C/C++, Python, software design, programming techniques..

  • Familiarity with and development experience on Windows and Linux.

  • Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.

  • Some travel is required for conferences and for on-site visits with external partners.

  • Strong problem-solving skills and the ability to work both independently and collaboratively in a fast-paced environment.

  • Excellent interpersonal and communication skills and a passion for keeping track with the latest advancements in AI technology.

Ways to stand out from the crowd:

  • Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, NVIDIA Model Optimizer.

  • Expertise with professional agentic AI use cases, i.e., digital content creation and productivity workflows.

  • Experience working with open-source LLM and GenAI software.

  • Detailed knowledge of the latest generation GPU architectures.

  • Experience with AI deployment on NPUs and ARM architectures.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative, autonomous and love a challenge, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 26, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA

About the company

NVIDIA

Large Enterprise

NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.