NVIDIA Corporation

Principal Software Engineer - Compute Infrastructure

NVIDIA Corporation$248K — $391K *
US-Anywhere
+ 2 other locationsRemote
Information Technology
11 - 15 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor’s degree in Engineering, Computer Science, Mathematics, or related field, or equivalent experience.
  • 15+ years in compute platform engineering, site reliability, or systems architecture with a focus on massive scale automation.
  • Expertise in Kubernetes architecture and deploying virtualization architectures (KubeVirt, OpenShift).
  • In-depth knowledge of hardware technologies like GPUs and high-speed networking.
  • Experience managing global environments across bare metal, virtual, and cloud infrastructures with GitOps posture.
  • Proficiency in programming languages such as Go/Python and infrastructure-as-code development (Terraform).
  • Strong leadership skills to influence technical direction in autonomous teams.

Responsibilities

  • Define and transform the global enterprise compute platform architecture.
  • Build operational foundations for internal AI inference platforms.
  • Develop strategies for capacity planning under hardware supply constraints.
  • Collaborate to drive cultural adoption of standard platforms within teams.
  • Lead migrations of large-scale legacy workloads to modern Kubernetes orchestration.

Benefits

  • Eligible for equity participation in the company.
  • Opportunity to work with cutting-edge technology in a hybrid environment.
  • Be part of a dynamic and innovative team focused on AI and infrastructure optimization.
  • Access to advanced tools and resources for professional development.
Full Job Description
We are seeking a highly skilled Principal Software Engineer to join our dynamic team. Our company is at the forefront of technological innovation, and we are dedicated to driving efficiency, defining platform architecture, and optimizing the performance of our infrastructure both on-prem and in the cloud. You will lead the architectural vision for a massive global platform and spearhead the operationalization of our internal frontier-class AI inference systems. Join us in this exciting endeavor! What You Will Be Doing: • Define Platform Architecture: Lead initiatives to architect and transform our global enterprise compute platform—running thousands of nodes and tens of thousands of VMs and containers via OpenShift and KubeVirt—by defining service tiers, SLAs, and automated cluster lifecycles. • Operationalize Frontier AI Infrastructure: Build the operational foundation for our internal AI inference platform scaling to frontier-class models. You will develop automated remediation pipelines, hardware watchdogs, and telemetry for pre-release, rack-scale GPU systems (including Blackwell and upcoming architectures). • Drive Strategic Capacity & Scale: Collect and review system data for capacity planning to navigate extreme hardware supply constraints. Develop proactive strategies, including public cloud bursting, hardware dogfooding, and evaluating alternative compute architectures (e.g., ARM). • Build the "Paved Road": Collaborate with highly autonomous NVIDIA engineering teams to drive cultural adoption of standard platforms. You will design compelling self-service architectures, APIs, and Terraform/OpenTofu providers that teams want to use. • Lead Complex Migrations: Evaluate existing application architectures and drive the fraught but critical migration of massive legacy workloads—including large-scale, long-running VDI environments—into modern Kubernetes orchestration. What We Need To See: • Bachelor’s degree in Engineering, Computer Science, Mathematics, or related field, or equivalent experience. • 15+ years of proven experience in compute platform engineering, site reliability, or systems architecture with a heavy focus on automation at massive scale. • Deep expertise in Kubernetes architecture and designing/deploying virtualization architectures, specifically operating VMs inside K8s (KubeVirt, OpenShift). • In-depth knowledge of hardware technologies (GPUs, high-speed backplane networking) with a track record of mitigating hardware-level failures, silent data corruption, and anomalies in large-scale environments. • Experience running large global environments spanning bare metal, virtualized infrastructure, and cloud with a unified GitOps posture (ArgoCD or similar). • Proficiency in programming languages such as Go and/or Python, alongside expert-level infrastructure-as-code development (Terraform, Config Management). • Strong leadership skills with the ability to influence technical direction across highly autonomous teams without relying on top-down mandates. Ways To Stand Out From The Crowd: • Hands-on experience managing bleeding-edge, pre-release hardware in production environments. • Deep understanding of advanced storage migrations and protocols (NFSv4, NVMe/TCP, Hyperconverged storage). • Solid understanding of microservices architecture and seamless multi-cloud deployment strategies (AWS, GCP). • Proven track record of building "Day 2" operational maturity (self-service, advanced auto-remediation, strict SLAs) from the ground up on existing foundations. #LI-Hybrid Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD. You will also be eligible for equity and . Applications for this job will be accepted at least until September 3, 2026. This posting is for an existing vacancy.  NVIDIA uses AI tools in its recruiting processes.

About NVIDIA Corporation

Nvidia, a global leader in graphics, gaming, and AI technology, offers Nvidia careers and internship opportunities for those passionate about driving innovation in the tech industry. you'll find a company committed to growth, teamwork, and leadership in computer science and machine learning domains.

About Nvidia

A Pioneer in Technology and Innovation

Nvidia has cemented its reputation as a powerhouse in developing advanced graphics processing units (GPUs) and has significantly contributed to the gaming industry's evolution. Moreover, its foray into AI and machine learning has opened new frontiers in technology, making Nvidia a beacon of innovation and a desirable workplace for ambitious tech professionals.

Job Opportunities

Diverse Positions in a Dynamic Field

Nvidia is continuously on the lookout for talented individuals across various domains, including hardware and software engineering, product design, marketing, and sales. Employment opportunities at Nvidia are vast, catering to a wide range of expertise and career aspirations.

Employment in Hardware and Graphics

For those fascinated by the intricacies of hardware and graphics technology, Nvidia offers positions that sit at the forefront of gaming and computing advancements.

Growth in Machine Learning and AI

Nvidia's leadership in AI and machine learning has created numerous vacancies for specialists eager to contribute to groundbreaking projects.

Recruitment in Computer Science

With the constant demand for innovation, Nvidia's recruitment efforts focus on computer science experts capable of pushing the boundaries of what's possible.

Internship Program

Opening Doors to Future Innovators

Nvidia's internship program is designed to nurture the next generation of technology leaders, offering hands-on experience in a culture that celebrates creativity and teamwork.

Benefits and Culture

Interns at Nvidia enjoy a plethora of benefits, from competitive stipends to mentorship opportunities, all within an environment that values growth and learning.

Opportunities for Students

Whether you're an undergraduate, a master's student, or a Ph.D. candidate, Nvidia's internships provide a real-world glimpse into the tech industry, offering valuable experience in various technology fields.

Pathways to Full-Time Employment

Many interns have transitioned into full-time positions, marking the start of successful careers at Nvidia. The internship program is more than a stepping stone into the company; it’s an investment in the professional development of interns. The goal is to ensure that interns are well-equipped for future challenges.

Nvidia Careers: More Than Just a Job

Nvidia offers more than just a job to its employees; it provides a front-row seat on the journey into the future of technology. Nvidia stands as a pillar of innovation with its vast opportunities in hardware, graphics, gaming, machine learning, and computer science. Nvidia careers serve as a launching pad for talented workers who aim to redefine the technological landscape. Whether through full-time positions or internships, joining Nvidia means contributing to a legacy of breakthroughs and becoming part of a global community dedicated to pushing the boundaries of what's possible.
Learn more about NVIDIA Corporation
Size
22,473 employees
Market Cap
$350.4 billion
Industry
Net Income
$4.3 billion
Founded
1993
5 Year Trend
+31.3%
Revenue
$16.6 billion
NASDAQ

Similar Jobs

More Jobs at NVIDIA Corporation

More Information Technology Jobs

Find similar Principal Software Engineer - Compute Infrastructure jobs: