NVIDIA

NVIDIA

Posted via Workday

Senior Cloud Operations Engineer

Apply by Aug 23, 2026

Posted Aug 19, 2026

Role at a glance

Salary
$184K – $287.5K/yr
Location
2 Locations, California, United States
Work arrangement
On-site
Employment
Full-time
Experience
8+ years of hands-on experience building/supporting complex services
Education
BS/MS in Computer Science (or equivalent experience)

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Senior Operations Engineer will join NVIDIA’s NGC Cloud team to improve the efficiency, reliability, and scalability of systems supporting global business operations. The role focuses on automation, monitoring, operational workflows, security, and coordination across engineering, operations, and security teams.

What You'll Do

  • Drive day-to-day interactions with NVIDIA-wide IT subsystems across infrastructure and applications
  • Craft and maintain GitLab CI/CD pipelines for build, test, and deployment workflows
  • Monitor system health, maintain dashboards, create alerts, and produce operational reports
  • Perform user offboarding, access reviews, and compliance-related tasks across multiple systems
  • Coordinate changes and releases between engineering, operations, and security teams
  • Enforce security guidelines, manage vulnerability remediation, and maintain operational documentation and SOPs

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Python automation; monitoring tools such as Prometheus, Grafana, Datadog, CloudWatch, and Splunk; ITSM practices; secure and compliant offboarding and access-related tasks; IT operations and system workflows; core Java including Collections API, Streams API, Concurrency, and I/O; RDBMS and NoSQL databases including Cassandra, DynamoDb, and Redis; cross-team communication and documentation.

Required

  • 8+ years of hands-on experience building/supporting complex services
  • BS/MS in Computer Science (or equivalent experience)
  • Knowledge in Python for automation, data handling, and tool development
  • Experience with monitoring tools such as Prometheus, Grafana, Datadog, CloudWatch, and Splunk
  • Familiarity with ITSM practices, including incident, problem, and modification processes
  • Ability to perform secure and compliant offboarding and access-related tasks
  • Strong understanding of IT operations and system workflows
  • Knowledge in core Java - Collections API, Streams API, Concurrency, I/O

Preferred

  • Experience designing or implementing automation pipelines or internal operational tools
  • Background in customer support, technical support, or customer-facing engineering roles
  • Prior work in a security-conscious or compliance-heavy environment
  • Ability to build end-to-end monitoring solutions, dashboards, and automated reporting
  • Strong documentation habits and a continuous-improvement approach

Original job description

Content provided by the employer

At NVIDIA, we are seeking a highly skilled Senior Operations Engineer to join our world-class NGC Cloud team. In this role, you will help drive the efficiency, reliability, and scalability of the systems that power our global business operations. This is an exceptional opportunity to shape how we automate, streamline, and support critical operational workflows across the organization. You will define how we implement innovative automation and support solutions, enabling teams to operate seamlessly and deliver impact at global scale—all within an encouraging and inclusive environment.

What you'll be doing:

  • Driving day-to-day interactions with NVIDIA wide IT subsystems, ensuring smooth operational workflows across infrastructure and applications.

  • Crafting and maintaining GitLab CI/CD pipelines to automate build, test, and deployment workflows.

  • Monitoring system health, building/maintaining dashboards, creating alerts, and producing operational reports.

  • Performing user offboarding, access reviews, and compliance-related tasks across multiple systems.

  • Drive interactions with various IT subsystems, ensuring API performance and integration stability meet defined SLAs and SLOs.

  • Coordinating changes and releases between engineering, operations, and security teams.

  • Enforcing security guidelines, managing vulnerability remediation, and collaborating with security teams on audits and assessments.

  • Maintaining documentation, SOPs, and process improvements to enhance operational maturity.

What we need to see:

  • 8+ years of hands-on experience building/supporting complex services and BS/MS in Computer Science (or equivalent experience).

  • Knowledge in Python for automation, data handling, and tool development.

  • Experience with monitoring tools (such as Prometheus, Grafana, Datadog, CloudWatch, Splunk) and reporting.

  • Familiarity with ITSM practices, including incident, problem, and modification processes.

  • Ability to perform secure and compliant offboarding and access-related tasks.

  • Strong understanding of IT operations and system workflows.

  • Knowledge in core Java - Collections API, Streams API, Concurrency, I/O.

  • Knowledge in RDBMS and NoSQL (Cassandra, DynamoDb, Redis) databases.

  • Excellent communication skills with the ability to collaborate across multiple teams.

  • Excellent documentation, problem-solving, and communication skills for cross-team alignment.

Ways to stand out from the crowd:

  • Experience designing or implementing automation pipelines or internal operational tools.

  • Background in customer support, technical support, or customer-facing engineering roles.

  • Prior work in a security-conscious or compliance-heavy environment.

  • Ability to build end-to-end monitoring solutions, dashboards, and automated reporting.

  • Strong documentation habits and a continuous-improvement approach.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 22, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA

About the company

NVIDIA

Large Enterprise

NVIDIA is a leading technology company renowned for its graphics processing units (GPUs) and innovative computing solutions that enhance visual experiences across multiple platforms, including gaming, scientific research, and artificial intelligence. Founded in 1993, the company has expanded its offerings to include powerful AI frameworks and deep learning platforms, making significant contributions to industries such as gaming, data centers, automotive, and healthcare. NVIDIA's commitment to pushing the boundaries of visual computing continues to drive advancements in both hardware and software, positioning the company at the forefront of emerging technologies and digital transformation.