Pinterest

Sr. Staff Software Engineer, Product ML Infrastructure

Pinterest • $245K — $429K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Proven ability to define technical strategy and implement infrastructure initiatives in complex settings.
  • Expertise in distributed machine learning systems and experience with both large-scale training and online inference.
  • Strong understanding of GPU performance aspects, including profiling, distributed execution, and optimization techniques.
  • Background in AI/ML modeling, particularly in recommender systems or Ads ranking.
  • Proficient in systems programming with C++, Java, or Python.
  • Demonstrates high ownership and sound judgment in reliability and operational excellence.
  • Experience utilizing AI for improving development speed and ensuring quality control.

Responsibilities

  • Establish the technical vision and roadmap for model training and serving within the PMLI team.
  • Architect solutions for distributed training, fine-tuning, and high-scale inference processes.
  • Enhance efficiencies in data loading and distributed execution related to GPU usage.
  • Create reliable platforms focusing on training and serving consistency.
  • Collaborate with AI/ML teams to safely integrate features at Pinterest scale.
  • Guide organization-wide architectural decisions and mentor senior engineers for elevated standards.
  • Leverage AI in development to boost prototyping speed while ensuring data integrity and correctness.

Benefits

  • Flexible work model promoting balance and collaboration.
  • Company culture that values equity, inclusion, and creativity.
  • Opportunities for professional development and career growth.
  • Transparent communication regarding salary and equity options.
Full Job Description
Sr. Staff Software Engineer, Product ML Infrastructure

Pinterest's Product ML Infrastructure (PMLI) team enables fast, safe, and efficient delivery of AI/ML solutions across Ads and Core critical products. We build unified data, training, feature, and inference infrastructure; this role will set technical direction across model training and serving, with a focus on GPU efficiency and large-scale ranking systems.

What you'll do:
  • Set the technical vision and roadmap for model training and serving across PMLI, with reusable interfaces to data and feature infrastructure.
  • Lead architectures for distributed training, fine-tuning, distillation, evaluation, and high-scale CPU/GPU inference.
  • Improve efficiency across data loading, distributed execution, GPU kernels and memory, compilation, quantization, scheduling, and capacity.
  • Build reliable, observable platforms with strong quality guarantees and training/serving consistency.
  • Partner with Ads and Core AI/ML teams to productionize features and models safely at Pinterest scale.
  • Drive cross-organizational architecture decisions, migrations, and operational standards; mentor senior engineers and raise the engineering bar.
  • Use AI-assisted development and analysis to accelerate prototyping, performance diagnosis, and validation while maintaining rigorous correctness and data safeguards.

What we're looking for:
  • A track record of setting technical strategy and delivering company-wide infrastructure initiatives in ambiguous environments.
  • Deep expertise in distributed ML systems, including production experience with both large-scale training and online inference.
  • Strong GPU performance knowledge, such as profiling, distributed execution, kernel and memory optimization, compilation, or quantization.
  • Experience with AI/ML modeling, recommender systems, Ads ranking, retrieval, feature platforms, or similarly demanding ML workloads.
  • Strong systems programming and design skills in C++, Java, or Python.
  • High ownership and sound judgment in reliability, security, cost, and operational excellence.
  • Demonstrated ability to use AI to improve speed and critically evaluate AI-assisted work, with accountability for correctness, quality, and sensitive data.
  • Bachelor's degree in Computer Science, Engineering, a related field, or equivalent experience.


Relocation Statement:
  • This position is not eligible for relocation assistance. Visit our PinFlex page to learn more about our working model.


In-Office Requirement Statement:
  • We recognize that the ideal environment for work is situational and may differ across departments. What this looks like day-to-day can vary based on the needs of each organization or role.
  • This role will need to be in the office for in-person collaboration 1-2 times per quarter and therefore needs to be within a commutable distance of our Palo Alto, CA or San Francisco, CA office.


#LI-REMOTE

#LI-AG8

At Pinterest we believe the workplace should be equitable, inclusive, and inspiring for every employee. In an effort to provide greater transparency, we are sharing the base salary range for this position. The position is also eligible for equity. Final salary is based on a number of factors including location, travel, relevant prior experience, or particular skills and expertise.

Information regarding the culture at Pinterest and benefits available for this position can be found here.

US based applicants only

$245,402-$429,454 USD

About Pinterest

Pinterest is a social media platform that allows users to discover and save ideas for recipes, home decor, fashion, and more. The company was founded in 2010 and is headquartered in San Francisco, California. Pinterest has over 400 million monthly active users and is available in over 30 languages. The company's mission is to help people discover and do what they love.
Learn more about Pinterest
Size
3,225 employees
Market Cap
$16 billion
Industry
Net Income
-$128.3 million
Founded
2009
5 Year Trend
+53.9%
Revenue
$1.6 billion
NASDAQ

Similar Jobs

More Jobs at Pinterest

More Information Technology Jobs

Find similar Sr. Staff Software Engineer, Product ML Infrastructure jobs: