Role at a glance
- Salary
- $124K – $241.5K/yr
- Location
- 3 Locations, California, United States
- Work arrangement
- On-site
- Employment
- Full-time
- Experience
- Relevant work or research experience.
- Education
- Pursuing or recently completed a Master’s or PhD degree (or equivalent experience) in Computer Science, Computer Engineering, or...
Spotted an issue?
We’ll check it against the original posting.
Role Summary
The AI Developer Technology Engineer will develop techniques to accelerate high-performance workloads for financial services and insurance-focused AI on NVIDIA CPUs and GPUs. The role focuses on parallel algorithms, GPU performance, complex AI and HPC workloads, and collaboration with research, hardware, compiler, tools, and broader developer communities.
What You'll Do
- Research, design, and develop techniques to accelerate high-performance workloads for FSI-focused AI on NVIDIA CPUs and GPUs.
- Analyze, optimize, and scale complex AI and HPC workloads for modern CPU and GPU architectures.
- Profile and eliminate performance bottlenecks across algorithms, kernels, and system-level behavior.
- Publish and present work in conferences, talks, and blogs.
- Collaborate with research, hardware, compiler, and tools teams to influence future hardware architectures, system software, libraries,...
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Relevant work or research experience; experience with low-level parallel programming such as CUDA; understanding of CPU/GPU architecture fundamentals and performance; fluency in C/C++; foundations in algorithms and software design; experience improving large-scale computational applications on GPUs; good understanding of linear algebra; strong communication and organization skills.
Required
- Pursuing or recently completed a Master’s or PhD degree (or equivalent experience) in Computer Science, Computer Engineering, or...
- Relevant work or research experience.
- Experience with low-level parallel programming (e.g., CUDA).
- Deep understanding of CPU/GPU architecture fundamentals and how they impact performance.
- Fluency in C/C++ and solid foundations in algorithms and software design.
- Experience improving the performance of large-scale computational applications on GPUs.
- Good understanding of linear algebra.
- Strong communication and organization skills, with a logical approach to problem solving and solid prioritization abilities.
Preferred
- Prior internship experience in a related field.
- Experience with inference optimization techniques and deploying optimized AI models in production.
- Experience with TensorRT, TensorRT-LLM, and cuTile.
- Background in capital markets with exposure to systematic/algorithmic strategies or quantitative trading.
- Experience parallelizing and optimizing machine learning methods such as decision trees, time series models, and Monte Carlo simulations.
- Knowledge of financial data models, pricing and risk simulation algorithms, portfolio optimization, or other finance-focused...
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Our work at NVIDIA is dedicated towards a computing model focused on visual and AI computing. For two decades, NVIDIA has pioneered visual computing, the art and science of computer graphics, with our invention of the GPU. The GPU has also shown to be spectacularly effective at solving some of the most complex problems in computer science. Today, NVIDIA’s GPU simulates human intelligence, running deep learning algorithms and acting as the brain of computers, robots and self-driving cars that can perceive and understand the world. We are looking to grow our company and teams with the smartest people in the world and there has never been a more exciting time to join our team!
We’re looking for an AI Developer Technology Engineer to push the limits of performance at the intersection of AI, high-performance computing, and financial markets. In this role, you’ll dive deep into parallel algorithms, GPUs, and complex systems to identify and eliminate bottlenecks, unlocking the full power of the world’s most advanced processing hardware. You’ll collaborate with top experts across industry and academia, influence next-generation platforms, and share your insights with the global developer community. Would you enjoy solving hard technical problems, love performance tuning, and want your work to have a visible impact across an entire industry? If so, we would love to invite you to consider this role!
What you will be doing:
Researching, designing, and developing groundbreaking techniques to accelerate high-performance workloads for FSI-focused, pioneering AI on NVIDIA CPUs and GPUs.
Working with leading technical experts to analyze, optimize, and scale complex AI and HPC workloads for modern CPU and GPU architectures.
Profiling and eliminating performance bottlenecks across the stack: from algorithms to kernels to system-level behavior.
Publishing and presenting your work in conferences, talks, and blogs to educate and inspire the broader developer community.
Influencing the design of future hardware architectures, system software, libraries, and programming models by collaborating closely with NVIDIA research, hardware, compiler, and tools teams.
What we need to see:
Pursuing or recently completed a Master’s or PhD degree (or equivalent experience) in Computer Science, Computer Engineering, or Electrical and Computer Engineering or related field.
Relevant work or research experience.
Experience with low-level parallel programming (e.g., CUDA).
Deep understanding of CPU/GPU architecture fundamentals and how they impact performance.
Fluency in C/C++ and solid foundations in algorithms and software design.
Experience improving the performance of large-scale computational applications on GPUs.
Good understanding of linear algebra.
Strong communication and organization skills, with a logical approach to problem solving and solid prioritization abilities.
Ways to stand out from the crowd:
Prior internship experience in a related field.
Experience with inference optimization techniques and deploying optimized AI models in production.
Experience with TensorRT, TensorRT-LLM, and cuTile.
Background in capital markets with exposure to systematic/algorithmic strategies or quantitative trading.
Experience parallelizing and optimizing machine learning methods such as decision trees, time series models, and Monte Carlo simulations as well as knowledge of financial data models, pricing and risk simulation algorithms, portfolio optimization, or other finance-focused applications and services.
NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 124,000 USD - 195,500 USD for Level 2, and 152,000 USD - 241,500 USD for Level 3.You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About the company
NVIDIA
Large Enterprise
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.