Role at a glance
- Salary
- $38 – $94/hr
- Location
- Santa Clara, California, United States
- Work arrangement
- On-site
- Employment
- Internship
- Experience
- Strong background in research with publications at top conferences.
- Education
- PhD
Spotted an issue?
We’ll check it against the original posting.
Role Summary
This Ph.D. internship focuses on researching and developing methods that advance large language and multimodal models within NVIDIA’s LLM research teams. The work is intended to transfer research into product groups and may result in prototypes, patents, products, or original research publications.
What You'll Do
- Research and develop novel methods for advancing large language and multimodal models.
- Collaborate with team members, other teams, and external researchers.
- Transfer research to product groups to enable new products or types of products.
- Deliver prototypes, patents, products, and/or publish original research.
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Active enrollment in a Ph.D. program in Computer Science, Electrical Engineering, or a related field for the full internship; anticipated graduation date must be indicated on a resume or CV. Programming experience may include Python, C++, CUDA, and deep learning frameworks such as PyTorch, TensorFlow, or JAX. Research experience may include large language and foundation models, transformer architectures, model efficiency and optimization, training and alignment, multimodal and vision-language models, or retrieval-augmented generation.
Required
- Python
- C++
- CUDA
- PyTorch, TensorFlow, JAX, or other deep learning frameworks
- Research publications at top conferences
- Excellent communication and collaboration skills
Preferred
- Experience with large-scale model training
Original job description
Content provided by the employer
Original job description
Content provided by the employer
By submitting your resume, you acknowledge that your Ph.D. Research Large Language Models internship application will be processed in accordance with NVIDIA’s Applicant Privacy Policy and you agree to our Terms of Service. We’ll review resumes on an ongoing basis, and a recruiter may reach out if your experience fits one of our many internship opportunities.
NVIDIA pioneered accelerated computing to tackle challenges no one else can solve. Our work in AI and digital twins is transforming the world's largest industries and profoundly impacting society — from gaming to robotics, self-driving cars to life-saving healthcare, climate change to virtual worlds where we can all connect and create.
Our internships offer an excellent opportunity to expand your career and get hands on experience with one of our industry leading LLM teams. We’re seeking strategic, ambitious, hard-working, collaborative, and creative individuals who are passionate about helping us tackle challenges no one else can solve.
Learn more about Research at NVIDIA.
What you will be doing:
- Research and develop novel methods for advancing the capabilities of large language and multimodal models.
- Collaborate with other team members, teams, and/or external researchers.
- Transfer your research to product groups to enable new products or types of products. Deliverable results include prototypes, patents, products, and/or publishing original research.
What we need to see:
- Must be actively enrolled in a university pursuing a Ph.D. degree in Computer Science, Electrical Engineering, or a related field, for the full duration of the internship; anticipated graduation date (month and year) must be clearly indicated on a resume or CV to be considered.
- Depending on the internship, prior experience or knowledge requirements could include the following programming skills and technologies:
Python, C++, CUDA, Deep Learning Framworks (PyTorch, Tensorflow, JAX, etc.)
- Strong background in research with publications at top conferences.
- Excellent communication and collaboration skills.
- Experience with large-scale model training is a plus.
Potential internships require research experience in at least one of the following areas:
- Large Language Models and Foundation Models
Transformer architectures
Knowledge distillation and data synthesis
Long-context methods
- Model Efficiency and Optimization
Model compression and pruning
Quantization
Inference optimization and acceleration
Parameter-efficient fine-tuning
Neural Architecture Search (NAS)
- Training and Alignment
Large-scale model training
Instruction tuning
Reinforcement Learning
Advanced reasoning and test-time inference
Few-shot and zero-shot learning
- Synthetic data generation
- Multimodal and Vision Language Models
- Retrieval-Augmented Generation (RAG)
Click here to learn more about NVIDIA, our early talent programs, benefits offered to students and other helpful student resources related to our latest technologies and endeavors.
Our internship hourly rates are a standard pay based on the position, your location, year in school, degree, and experience. The hourly rate for our interns is 38 USD - 94 USD.You will also be eligible for Intern benefits.
Applications are accepted on an ongoing basis.This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About the company
NVIDIA
Large Enterprise
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.
NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.