Senior Full Stack Software Engineer - DGX Cloud

NVIDIA Corporation • $152K — $287K *

US-Anywhere

+ 5 other locationsRemote

Consumer Technology

5 - 7 years of experience

3 weeks ago

By clicking Apply, I agree with Ladders' Terms of Use and Privacy Policy

Job Overview by Ladders

Qualifications

5+ years in a software engineering role with impactful contributions
Experience in large-scale production systems
BS in Computer Science or Engineering or equivalent experience
6+ years of full-stack engineering experience
3+ years building and shipping consumer-facing products
Proficient in React, TypeScript/JavaScript, and Golang
Experienced with SQL databases

Responsibilities

Participate in DGX Cloud team for GPU cluster production systems
Design and develop a scalable platform for GPU asset management
Collaborate across teams to ensure reliability and performance of AI clusters
Evaluate system failures and enhance services through incident management
Work with diverse technologies including React, Web Components, Typescript, Golang, PostgreSQL, Temporal, Bazel, and Kubernetes

Benefits

Eligible for equity
Comprehensive benefits package
Access to AI tools in recruiting processes
Professional development opportunities
Collaboration with a talented and innovative team

Full Job Description

We expect you to have significant software engineering experience with cluster operations, operator development, node health monitoring and working with GPU resource scheduling. We welcome out-of-the-box thinkers who can provide new ideas with strong execution bias. Expect to be constantly challenged, improving, and evolving for the better. You will help advance NVIDIA's capacity to build and deploy leading infrastructure solutions for a broad range of AI-based applications. If you're creative, passionate about GPUs, and love having fun, please apply today!

What you will be doing:

You will be part of an DGX Cloud team responsible for production systems that enable large scalable GPU clusters to be used for a variety of AI workloads.
Designing and developing a massively distributed scalable platform which would be used to identify, diagnose and remediate non-performant GPU assets.
Working with teams across NVIDIA to ensure production AI clusters run reliability and consistently with maximum performance. Evaluating system failures and improving services based on a well-defined incident management process.
Working across all of our product stack: React, Web Components, TypeScript, Golang, PostgreSQL, Temporal, Bazel, Kubernetes

What we need to see:

Direct experience in a software engineering role within a highly technical organization with demonstrable impact from your work.
Highly motivated with strong communication skills, you can work successfully with multi-functional teams, principles, and architects and coordinate effectively across organizational boundaries and geographies.
5+ years in similar role and experience on large-scale production systems. Experience with common software engineering principles, tools and techniques
You possess a BS in Computer Science or Engineering or equivalent experience
6+ years of experience doing full-stack engineering
3+ years building and shipping consumer-facing products
Proficiency in React, TypeScript/JavaScript, and Golang
Proficiency with a SQL database

Ways to stand out from the crowd:

Technical competency in managing and automating large-scale distributed systems independent of cloud providers. Advanced hands-on experience and deep understanding of cluster management systems (Kubernetes, Slurm, Base Command Manager).
Empathy for users, attention to detail, and a passion for creating world-class user experiences
Prior experience in asynchronous workflows and/or event driven architecture.
Proven operational excellence in maintaining reliable and performant infrastructure.
A good understanding of how to use LLMs responsibly and the perils of blindly consuming their output

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until May 17, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

About NVIDIA Corporation

Nvidia, a global leader in graphics, gaming, and AI technology, offers Nvidia careers and internship opportunities for those passionate about driving innovation in the tech industry. you'll find a company committed to growth, teamwork, and leadership in computer science and machine learning domains.

About Nvidia

A Pioneer in Technology and Innovation

Nvidia has cemented its reputation as a powerhouse in developing advanced graphics processing units (GPUs) and has significantly contributed to the gaming industry's evolution. Moreover, its foray into AI and machine learning has opened new frontiers in technology, making Nvidia a beacon of innovation and a desirable workplace for ambitious tech professionals.

Job Opportunities

Diverse Positions in a Dynamic Field

Nvidia is continuously on the lookout for talented individuals across various domains, including hardware and software engineering, product design, marketing, and sales. Employment opportunities at Nvidia are vast, catering to a wide range of expertise and career aspirations.

Employment in Hardware and Graphics

For those fascinated by the intricacies of hardware and graphics technology, Nvidia offers positions that sit at the forefront of gaming and computing advancements.

Growth in Machine Learning and AI

Nvidia's leadership in AI and machine learning has created numerous vacancies for specialists eager to contribute to groundbreaking projects.

Recruitment in Computer Science

With the constant demand for innovation, Nvidia's recruitment efforts focus on computer science experts capable of pushing the boundaries of what's possible.

Internship Program

Opening Doors to Future Innovators

Nvidia's internship program is designed to nurture the next generation of technology leaders, offering hands-on experience in a culture that celebrates creativity and teamwork.

Benefits and Culture

Interns at Nvidia enjoy a plethora of benefits, from competitive stipends to mentorship opportunities, all within an environment that values growth and learning.

Opportunities for Students

Whether you're an undergraduate, a master's student, or a Ph.D. candidate, Nvidia's internships provide a real-world glimpse into the tech industry, offering valuable experience in various technology fields.

Pathways to Full-Time Employment

Many interns have transitioned into full-time positions, marking the start of successful careers at Nvidia. The internship program is more than a stepping stone into the company; it’s an investment in the professional development of interns. The goal is to ensure that interns are well-equipped for future challenges.

Nvidia Careers: More Than Just a Job

Nvidia offers more than just a job to its employees; it provides a front-row seat on the journey into the future of technology. Nvidia stands as a pillar of innovation with its vast opportunities in hardware, graphics, gaming, machine learning, and computer science. Nvidia careers serve as a launching pad for talented workers who aim to redefine the technological landscape. Whether through full-time positions or internships, joining Nvidia means contributing to a legacy of breakthroughs and becoming part of a global community dedicated to pushing the boundaries of what's possible.

Learn more about NVIDIA Corporation

Size

22,473 employees

Market Cap

$350.4 billion

Industry

Manufacturing & Automotive

Net Income

$4.3 billion

Founded

1993

5 Year Trend

+31.3%

Revenue

$16.6 billion

NASDAQ

NVDA

* Ladders Estimates

Similar Jobs

Software Engineer
$160K — $190K *
Osaro
San Francisco, CA 94112 (San Francisco County)
Reposted Today
Senior APEX Developer (Top Secret Clearance Required)
$180K — $260K *
Contact Government Services, LLC
San Francisco, CA 94112 (San Francisco County)
Reposted Today
Senior Software Engineer
$129K — $193K *
The MITRE Corporation
El Segundo, CA 90245 (Los Angeles County)
Today
Senior Software Engineer, Consensus
$180K — $300K *
Anza
New York, NY 10025 (New York County)
Reposted Today
Sr. Software Development Engineer, Aurora Storage
$168K — $227K *
Amazon
Redmond, WA 98052 (King County)
Reposted Today
Senior Software Engineer, FTVX Team (Whole World)
$168K — $227K *
Amazon
Seattle, WA 98115 (King County)
Reposted Today

Get Ready For Your
Next Interview

More Jobs at NVIDIA Corporation

Customer Technical Program Manager
$168K — $258K *
Austin, TX 78745 (Travis County)
Yesterday
Consumer Technology
In-Person
Customer Technical Program Manager
$168K — $258K *
Santa Clara, CA 95051 (Santa Clara County)
Yesterday
Consumer Technology
In-Person
Senior Software Engineer, DGX Cloud AI Infrastructure
$184K — $356K *
Redmond, WA 98052 (King County)
Yesterday
Enterprise Technology
In-Person
Senior Software Engineer, DGX Cloud AI Infrastructure
$184K — $356K *
Remote
Yesterday
Information Technology
Remote in United States
Senior AI Infrastructure Engineer - DGX Cloud
$152K — $287K *
Santa Clara, CA 95051 (Santa Clara County)
Yesterday
Enterprise Technology
In-Person

More Consumer Technology Jobs

Worldwide Readiness Product Launch Analyst
$100K — $130K *
Apple
Cupertino, CA 95014 (Santa Clara County)
Reposted Today
Analog IC Design Engineer
$130K — $180K *
Apple
Cupertino, CA 95014 (Santa Clara County)
Reposted Today
Design Engineer
$80K — $120K *
vvd
Remote
Reposted Today
GPU Silicon Prototype Engineer
$120K — $150K *
Apple
Austin, TX 78745 (Travis County)
Reposted Today
Back-End Software Engineer
$90K — $130K *
Bask Health
Remote
Reposted Today

Find similar Senior Full Stack Software Engineer - DGX Cloud jobs:

Nationwide Remote