Lead Site Reliability Engineer, Platforms

Zoom Video Communications, Inc.

$124K — $271K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years of SRE or DevOps experience with production infrastructure at scale
  • Proficient in at least one programming language beyond scripting (e.g., Python, Go, Java)
  • Experience deploying and managing CI/CD pipelines using tools like Git, Jenkins, or Argo CD
  • Skilled in operating cloud infrastructure using Terraform and Kubernetes
  • Experience with observability tools such as ELK, Prometheus, or Grafana
  • Strong communication skills to relay complex technical concepts to diverse audiences
  • US citizenship or Green Card status required

Responsibilities

  • Design and scale DevOps platform services focusing on Kubernetes and cloud systems
  • Define technical roadmaps for infrastructure automation and security
  • Collaborate with service teams to address platform needs and deliver solutions
  • Advocate for SRE best practices including infrastructure as code and monitoring
  • Mentor team members through system design and deployment

Benefits

  • Variety of perks to support physical, mental, emotional, and financial health
  • Options that promote work-life balance
  • Opportunities to contribute to community initiatives
  • Award-winning workplace culture
  • Hybrid work environment accommodation
Full Job Description
What You Can Expect

As a Lead Staff Site Reliability Engineer, you will be one of the technical leads for our DevOps Platforms organization. This group is responsible for DevOps Platforms including cloud infrastructure, physical data center orchestration, critical security services, and our Zoom for Government (ZfG) environment. You will be an uber tech lead working across a broad area, defining projects and guiding work across various teams. Your scope of work is wide and you will have the opportunity to improve our datacenter kubernetes infrastructure, our cloud infrastructure, our security posture, and our operation of ZfG environments. Broadly speaking, you are an exemplary SRE and you will guide all of our teams toward SRE best practices (automation, monitoring, infrastructure as code, etc).

Responsibilities
  • Design and scale DevOps platform services including Kubernetes infrastructure, cloud systems, and compliance-ready environments
  • Define technical roadmaps and architectural direction for infrastructure automation and security
  • Partner with service teams to understand platform needs and deliver solutions that improve reliability and efficiency
  • Establish and advocate for SRE best practices including infrastructure as code, monitoring, and incident management
  • Mentor team members through design, implementation, and production deployment of complex systems


What We're Looking For
  • Bring 8+ years of SRE or DevOps experience building and operating production infrastructure at scale
  • Code proficiently in at least one programming language beyond scripting (e.g., Python, Go, Java)
  • Deploy and manage CI/CD pipelines using tools like Git, Jenkins, Argo CD, or JFrog
  • Operate cloud infrastructure on AWS, OCI, or similar platforms using Terraform and Kubernetes
  • Implement observability solutions with logging and monitoring tools such as ELK, Prometheus, or Grafana
  • Communicate complex technical concepts clearly to diverse audiences including security teams, senior leadership, and external auditors
  • Participate in on-call rotations and lead incident response to maintain system reliability
  • Hold a degree in Computer Science or related field, or equivalent practical experience
  • Hold US citizenship, or Greencard status


Preferred

  • Have experience with security from an SRE perspective
  • Have experience with Identity security (e.g. IAM, workload identity, zero trust) and tools (e.g. Teleport, Okta)
  • Have experience operating Government environments and understanding their compliance requirements
  • Have experience with system design and distributed computing at scale
  • Ability to speak Chinese/Mandarin is a plus, but not required


Salary Range or On Target Earnings:

Minimum:
$124,000.00

Maximum:
$271,200.00

In addition to the base salary and/or OTE listed Zoom has a Total Direct Compensation philosophy that takes into consideration; base salary, bonus and equity value.

Note: Starting pay will be based on a number of factors and commensurate with qualifications & experience.

We also have a location based compensation structure; there may be a different range for candidates in this and other locations

At Zoom, we offer a window of at least 5 days for you to apply because we believe in giving you every opportunity. Below is the potential closing date, just in case you want to mark it on your calendar. We look forward to receiving your application!

Anticipated Position Close Date:

08/31/26

Ways of Working
Our structured hybrid approach is centered around our offices and remote work environments. The work style of each role, Hybrid, Remote, or In-Person is indicated in the job description/posting.

Benefits
As part of our award-winning workplace culture and commitment to delivering happiness, our benefits program offers a variety of perks, benefits, and options to help employees maintain their physical, mental, emotional, and financial health; support work-life balance; and contribute to their community in meaningful ways. Click Learn for more information.

Similar Jobs

More Jobs at Zoom Video Communications, Inc.

More Information Technology Jobs

Find similar Lead Site Reliability Engineer, Platforms jobs: