NVIDIA

NVIDIA

Posted via Workday

Senior Solutions Architect, Robotics Foundation Model Training

Posted Aug 28, 2026

Role at a glance

Job function
Software Engineering & IT Solutions Architecture
Salary
$152K – $287.5K/yr
Location
Santa Clara, California, United States
Work arrangement
On-site
Employment
Full-time
Experience
5+ years of industry or research experience in deep learning, distributed computing, or large-scale model training.
Education
MS, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Electrical or Computer Engineering, Robotics, or a...

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Applied Engineer will help scale robotics foundation model training from experimentation to production as part of a Physical AI team spanning data generation, multimodal model training, robotics simulation, and deployment. The role collaborates with researchers, ML engineers, product and engineering teams, and customer teams to improve robotics model workflows and influence Physical AI platform development.

What You'll Do

  • Architect and optimize end-to-end training workflows for robotics foundation models.
  • Build proof-of-concepts, reference architectures, and agentic workflows for experimentation, benchmarking, and model improvement.
  • Scale pre-training, fine-tuning, and reinforcement learning workloads across multi-GPU and multi-node systems.
  • Identify and eliminate data pipeline bottlenecks across storage, networking, preprocessing, and data loading for multimodal datasets.
  • Provide feedback to NVIDIA product and engineering teams to shape future Physical AI platforms.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Hands-on experience training or optimizing multimodal or foundation models; experience across pre-training, supervised fine-tuning, RL or other post-training methods, evaluation, and model optimization; strong expertise in distributed training techniques on multi-GPU or multi-node systems; expertise with PyTorch, NVIDIA NeMo, JAX, or Hugging Face Transformers; experience building or working with high-throughput data pipelines for large-scale training; strong communication skills.

Required

  • MS, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Electrical or Computer Engineering, Robotics, or a...
  • 5+ years of industry or research experience in deep learning, distributed computing, or large-scale model training.
  • Hands-on experience training or optimizing multimodal or foundation models.
  • Experience across pre-training, supervised fine-tuning, RL or other post-training methods, evaluation, and model optimization.
  • Strong expertise in distributed training techniques on multi-GPU or multi-node systems.
  • Expertise with PyTorch, NVIDIA NeMo, JAX, or Hugging Face Transformers.
  • Experience building or working with high-throughput data pipelines for large-scale training.
  • Strong communication skills with the ability to effectively collaborate across Researchers, Engineers and executives.

Preferred

  • Familiarity with NVIDIA AI and robotics platforms.
  • Experience with robotics AI workloads, including reinforcement learning in simulation and synthetic data generation.
  • Experience profiling and optimizing workloads using tools such as Nsight Systems, Nsight Compute, or PyTorch Profiler.
  • Demonstrated impact improving training efficiency and scaling performance.
  • Experience building agentic workflows for automated experimentation, model evaluation, data analysis, or research acceleration.

Original job description

Content provided by the employer

We are building a team of innovators to help partners develop and adopt the next generation of Physical AI, spanning data generation, large-scale multimodal model training, robotics simulation and deployment!

We are looking for a hands-on Applied Engineer with deep expertise in training foundation models at scale and a strong background in robotics. This role operates at the intersection of innovative AI research, accelerated computing and real-world applications, offering a unique opportunity to work directly with model builders to scale cutting edge Robotics Models from experimentation to production. Collaboration spans research, engineering, and customer teams, influencing both product direction and applied AI adoption. Come join us and help shape the future of robotics foundation model training!

What You’ll Be Doing:

  • Engage with Researchers and ML engineers to architect and optimize end-to-end training workflows for robotics foundation models, like World Models, VLAs, WAMs.

  • Build proof-of-concepts, reference architectures, and agentic workflows that accelerate experimentation, benchmarking, and model improvement of NVIDIA’s Robotics Open model platforms like Cosmos and GR00T.

  • Scale pre-training, fine-tuning, and reinforcement learning workloads across multi-GPU and multi-node systems, improving utilization, throughput, and memory efficiency.

  • Identify and eliminate data pipeline bottlenecks across storage, networking, preprocessing, and data loading for multimodal datasets (video, sensor data, trajectories).

  • Collaborate with NVIDIA product and engineering teams to provide feedback that shapes future Physical AI platforms

What We Need to See:

  • MS, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Electrical or Computer Engineering, Robotics, or a related field.

  • 5+ years of industry or research experience in deep learning, distributed computing, or large-scale model training.

  • Hands-on experience training or optimizing multimodal or foundation models (e.g., VLMs, VLAs, World Models), ideally in robotics settings.

  • Experience across the AI model lifecycle, including pre-training, supervised fine-tuning, RL or other post-training methods, evaluation, and model optimization.

  • Strong expertise in distributed training techniques (data/model/pipeline parallelism, sharding, check-pointing) on multi-GPU or multi-node systems.

  • Expertise with multimodal training frameworks such as PyTorch, NVIDIA NeMo, JAX, or Hugging Face Transformers.

  • Experience building or working with high-throughput data pipelines for large-scale training, including storage bandwidth, network throughput, and preprocessing (e.g., decoding, tokenization, batching)

  • Strong communication skills with the ability to effectively collaborate across Researchers, Engineers and executives.

Ways to Stand Out From the Crowd:

  • Familiarity with NVIDIA AI and robotics platforms (e.g., Cosmos, GR00T, NeMo, Isaac Sim, Isaac Lab)

  • Experience with robotics AI workloads, including reinforcement learning in simulation and synthetic data generation.

  • Experience profiling and optimizing workloads using tools such as Nsight Systems, Nsight Compute, or PyTorch Profiler

  • Demonstrated impact improving training efficiency and scaling performance

  • Experience building agentic workflows for automated experimentation, model evaluation, data analysis, or research acceleration

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 31, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA

About the company

NVIDIA

Large Enterprise

NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.