NVIDIA Corporation

Senior Software Engineer - Local AI

NVIDIA Corporation$152K — $287K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's, Master's, or PhD in Computer Science, Software Engineering, Mathematics, or related field (or equivalent experience).
  • Excellent C++ programming and debugging skills with strong understanding of data structures and algorithms.
  • 5+ years of experience in AI inferencing pipelines with frameworks like ONNX RT, PyTorch, Tensor RT, llama.cpp, and vLLM.
  • Strong analytical and problem-solving abilities with multitasking effectiveness in dynamic environments.
  • Outstanding written and oral communication skills for collaborating with teams.

Responsibilities

  • Partner with NVIDIA's software, research, architecture, and product teams to align strategies for AI on RTX and DGX PCs.
  • Collaborate with industry partners to drive innovation in AI across domains like graphics and edge devices.
  • Conduct in-depth analysis and optimize AI models and data processing pipelines for performance on GPUs.
  • Identify and implement optimization techniques such as quantization, distillation, and pruning for large AI models.
  • Fine-tune and compress AI models to be suitable for edge devices.

Benefits

  • Equity opportunities for employees.
  • Eligibility for various benefits.
Full Job Description
There is a growing emphasis on running AI models locally, closer to the source of data. This approach reduces latency, improves real-time processing, and addresses privacy concerns by minimizing the need for sending data to centralized servers. As technology continues to advance, we can expect client-side AI (local execution) to play a pivotal role in crafting the digital landscape.

Local AI seeks a Senior Systems Software Engineer interested in solving client-side AI challenges on Windows and Linux PCs with limited resources.

What You'll Be Doing:
  • Partnering with NVIDIA software, research, architecture, and product teams to align strategies and technical needs for fostering the ecosystem of AI on RTX and DGX PCs.
  • Collaborate closely with industry partners to advance AI across critical domains-including graphics, web browsers, and edge devices-by driving innovation in both open and closed source technologies with emphasis on system level support.
  • Improving performance on current and next-generation GPU architectures by conducting in-depth analysis and end-to-end optimization of AI models, data processing pipelines, and inference runtime features.
  • Identifying, evaluating, and implementing compute and memory optimization techniques-such as quantization, distillation, and pruning-for large AI models; fine-tuning and compressing models to fit edge devices.


What We Need to See:
  • Bachelor's, Master's, or PhD in Computer Science, Software Engineering, Mathematics, or a related field (or equivalent experience).
  • Excellent C++ programming and debugging skills with a strong understanding of data structures and algorithms.
  • 5+ years of experience with proficiency in AI inferencing pipelines and applications using ML/DL frameworks, including ONNX RT, PyTorch, Tensor RT, llama.cpp and vLLM.
  • Strong analytical and problem-solving abilities, with the ability to multitask effectively in a dynamic environment.
  • Outstanding written and oral communication skills enabling effective collaboration with management and engineering teams.


Ways To Stand Out from The Crowd:
  • Understanding modern techniques in Machine Learning, Deep Neural Networks, and Generative AI with relevant contributions to major open-source projects will be a plus.
  • Consistent track record of delivering end-to-end products with geographically distributed teams in multinational product companies.
  • Proficiency in lower-level system/GPU programming, CUDA, developing high-performance systems.
  • Hands-on experience with building applications using APIs like ONNX RT, DirectX, PyTorch, TensorRT, Vulkan, llama.cpp.


Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 1, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

About NVIDIA Corporation

Nvidia, a global leader in graphics, gaming, and AI technology, offers Nvidia careers and internship opportunities for those passionate about driving innovation in the tech industry. you'll find a company committed to growth, teamwork, and leadership in computer science and machine learning domains.

About Nvidia

A Pioneer in Technology and Innovation

Nvidia has cemented its reputation as a powerhouse in developing advanced graphics processing units (GPUs) and has significantly contributed to the gaming industry's evolution. Moreover, its foray into AI and machine learning has opened new frontiers in technology, making Nvidia a beacon of innovation and a desirable workplace for ambitious tech professionals.

Job Opportunities

Diverse Positions in a Dynamic Field

Nvidia is continuously on the lookout for talented individuals across various domains, including hardware and software engineering, product design, marketing, and sales. Employment opportunities at Nvidia are vast, catering to a wide range of expertise and career aspirations.

Employment in Hardware and Graphics

For those fascinated by the intricacies of hardware and graphics technology, Nvidia offers positions that sit at the forefront of gaming and computing advancements.

Growth in Machine Learning and AI

Nvidia's leadership in AI and machine learning has created numerous vacancies for specialists eager to contribute to groundbreaking projects.

Recruitment in Computer Science

With the constant demand for innovation, Nvidia's recruitment efforts focus on computer science experts capable of pushing the boundaries of what's possible.

Internship Program

Opening Doors to Future Innovators

Nvidia's internship program is designed to nurture the next generation of technology leaders, offering hands-on experience in a culture that celebrates creativity and teamwork.

Benefits and Culture

Interns at Nvidia enjoy a plethora of benefits, from competitive stipends to mentorship opportunities, all within an environment that values growth and learning.

Opportunities for Students

Whether you're an undergraduate, a master's student, or a Ph.D. candidate, Nvidia's internships provide a real-world glimpse into the tech industry, offering valuable experience in various technology fields.

Pathways to Full-Time Employment

Many interns have transitioned into full-time positions, marking the start of successful careers at Nvidia. The internship program is more than a stepping stone into the company; it’s an investment in the professional development of interns. The goal is to ensure that interns are well-equipped for future challenges.

Nvidia Careers: More Than Just a Job

Nvidia offers more than just a job to its employees; it provides a front-row seat on the journey into the future of technology. Nvidia stands as a pillar of innovation with its vast opportunities in hardware, graphics, gaming, machine learning, and computer science. Nvidia careers serve as a launching pad for talented workers who aim to redefine the technological landscape. Whether through full-time positions or internships, joining Nvidia means contributing to a legacy of breakthroughs and becoming part of a global community dedicated to pushing the boundaries of what's possible.
Learn more about NVIDIA Corporation
Size
22,473 employees
Market Cap
$350.4 billion
Industry
Net Income
$4.3 billion
Founded
1993
5 Year Trend
+31.3%
Revenue
$16.6 billion
NASDAQ

Similar Jobs

More Jobs at NVIDIA Corporation

More Information Technology Jobs

Find similar Senior Software Engineer - Local AI jobs: