Pinterest

Sr. Site Reliability Engineer, tvScientific

Pinterest$139K — $287K *
US-AnywhereRemote in San Francisco, CA
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 4+ years in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
  • Strong hands-on AWS experience in production environments
  • Deep expertise in Kubernetes operations and troubleshooting
  • Proven multi-tenancy management in Kubernetes
  • Experience with ArgoCD and GitOps workflows
  • Strong skills in Terraform/Terragrunt for infrastructure management
  • Solid scripting skills in Bash and/or Python
  • Experience with CI/CD pipelines, particularly GitHub Actions
  • Strong troubleshooting skills across Linux and distributed systems
  • Ability to validate AI-assisted processes and maintain accountability

Responsibilities

  • Ensure the reliability, availability, and performance of production infrastructure
  • Operate and scale Kubernetes platforms for multi-tenant workloads
  • Manage GitOps-based deployment workflows using ArgoCD and Helm
  • Drive infrastructure provisioning and change management with Terraform/Terragrunt
  • Build and support CI/CD automation using GitHub Actions
  • Lead incident response and post-incident improvement efforts
  • Reduce operational toil through scripting and process automation
  • Enhance observability practices across systems
  • Support secure operations and platform guardrails
  • Collaborate with application, security, and platform teams

Benefits

  • Flexible in-office requirements based on departmental needs
  • Equity options available for employees
  • Inclusivity and equity focus in the workplace
  • Transparency in salary structure
  • Ability to work in various situational environments based on role requirements
Full Job Description
About tvScientific

tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. We leverage massive data and cutting-edge science to automate and optimize TV advertising to drive business outcomes. Our solution combines media buying, optimization, measurement, and attribution in one, efficient platform. Our platform is built by industry leaders with a long history in programmatic advertising, digital media, and ad verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.

We are seeking a Senior Site ReliabilityEngineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows. This role will be instrumental in advancing the reliability, scalability, automation, observability, and operational maturity of our infrastructure and delivery ecosystem.

The ideal candidate is a highly hands-on engineer with strong production experience and a proven ability to build and support resilient platforms using infrastructure as code, automation, and modern Kubernetes operational practices.

What you'll do:
  • Ensuring the reliability, availability, and performance of production infrastructure and platform services
  • Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads
  • Managing GitOps-based deployment workflows using ArgoCD and Helm
  • Driving infrastructure provisioning and change management through Terraform/Terragrunt
  • Building and supporting CI/CD automation and deployment workflows using GitHub Actions
  • Leading incident response efforts, root cause analysis, and post-incident improvement initiatives
  • Reducing operational toil through scripting, tooling, and process automation
  • Advancing observability practices across logs, metrics, traces, dashboards, and alerting
  • Supporting secure secrets integration, IAM-aware operations, and platform guardrails
  • Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes


What we're looking for:
  • 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
  • Strong hands-on experience operating AWS in production environments
  • Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration
  • Proven experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns
  • Experience implementing and operating ArgoCD within a GitOps delivery model
  • Strong hands-on experience with Helm
  • Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management
  • Solid scripting and automation skills using Bash and/or Python
  • Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions
  • Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems
  • Experience with monitoring, alerting, and observability in production environments
  • Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow-through after outages
  • Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams
  • Bachelor's degree in computer science, engineering, a related field or equivalent experience
  • Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs
  • Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review)
  • High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables.


In-Office Requirement Statement:
  • We recognize that the ideal environment for work is situational and may differ across departments. What this looks like day-to-day can vary based on the needs of each organization or role.

Relocation Statement:
  • This position is not eligible for relocation assistance. Visit our PinFlex page to learn more about our working model.

#LI-SM4

#LI-REMOTE

At Pinterest we believe the workplace should be equitable, inclusive, and inspiring for every employee. In an effort to provide greater transparency, we are sharing the base salary range for this position. The position is also eligible for equity. Final salary is based on a number of factors including location, travel, relevant prior experience, or particular skills and expertise.

Information regarding the culture at Pinterest and benefits available for this position can be found here.

US based applicants only

$139,764-$287,749 USD

About Pinterest

Pinterest is a social media platform that allows users to discover and save ideas for recipes, home decor, fashion, and more. The company was founded in 2010 and is headquartered in San Francisco, California. Pinterest has over 400 million monthly active users and is available in over 30 languages. The company's mission is to help people discover and do what they love.
Learn more about Pinterest
Size
3,225 employees
Market Cap
$16 billion
Industry
Net Income
-$128.3 million
Founded
2009
5 Year Trend
+53.9%
Revenue
$1.6 billion
NASDAQ

Similar Jobs

More Jobs at Pinterest

More Information Technology Jobs

Find similar Sr. Site Reliability Engineer, tvScientific jobs: