Site Reliability Engineer III, GCP

Optimum

• $133K — $219K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years of experience in SRE, Platform Engineering, Cloud Infrastructure, DevOps, or Network Engineering.
  • 5+ years of hands-on experience with production workloads on Google Cloud Platform.
  • Strong expertise in GCP, GKE, Kubernetes, and Terraform.
  • Experience with cloud networking including NCC, Interconnect, and Cloud VPN.
  • Proven background supporting highly available production environments.
  • Familiarity with SLOs, incident management, and automation practices.
  • Proficient in scripting languages like Python or Bash.

Responsibilities

  • Design and operate enterprise-scale GCP infrastructure and platforms.
  • Build and manage solutions using GKE, Cloud Run, and Compute Engine.
  • Architect cloud networking solutions including Cloud VPN and Shared VPC.
  • Develop Infrastructure as Code using Terraform and automate workflows.
  • Define and manage SLOs while improving observability through automation.
  • Lead incident response and reliability improvement initiatives.
  • Implement cloud security and operational best practices.

Benefits

  • Mentorship opportunities to grow into leadership roles.
  • Chance to work on cutting-edge cloud technologies in a dynamic environment.
  • Access to ongoing training and professional development.
  • Collaborative workplace promoting innovation and excellence.
Full Job Description
Job Summary

Optimum is seeking a Lead Site Reliability Engineer (SRE) to design, build, and operate secure, scalable, and highly available platforms on Google Cloud Platform (GCP). This hands-on engineering role combines cloud infrastructure, networking, automation, Kubernetes, and reliability engineering to support critical enterprise and customer-facing applications.

The ideal candidate brings deep expertise in GCP, cloud networking, Infrastructure as Code, observability, and incident management, along with a passion for improving reliability through automation and engineering excellence. You are a hands-on engineer with deep GCP and cloud networking expertise who enjoys solving complex operational challenges through automation, platform engineering, and reliability-focused design. You combine technical depth with leadership to build resilient, scalable cloud platforms that enable teams to deliver software with confidence.

Responsibilities

  • Design and operate enterprise-scale GCP infrastructure and platforms.
  • Build and manage solutions using GKE, Cloud Run, Compute Engine, Cloud SQL, Pub/Sub, Cloud Storage, and Cloud Monitoring.
  • Architect and support cloud networking, including Network Connectivity Center (NCC), Interconnect, Cloud VPN, Cloud Router, Shared VPC, and Private Service Connect.
  • Develop Infrastructure as Code using Terraform and automate deployment and operational workflows through CI/CD pipelines.
  • Define and manage SLOs, improve observability, and reduce operational toil through automation.
  • Lead incident response, root cause analysis, and reliability improvement initiatives.
  • Implement cloud security, IAM, governance, and operational best practices.
  • Partner with application, network, and security teams to deliver reliable cloud-native platforms.
  • Mentor engineers and help establish platform and SRE standards across the organization.


Qualifications

  • 8+ years of experience in SRE, Platform Engineering, Cloud Infrastructure, DevOps, or Network Engineering.
  • 5+ years of hands-on experience operating production workloads on Google Cloud Platform.
  • Strong expertise in:
    • GCP and GKE
    • Kubernetes
    • Terraform
    • Cloud Networking
    • IAM
    • Monitoring and Observability
    • Infrastructure as Code
  • Experience with NCC, Interconnect, Cloud VPN, Cloud Router, Shared VPC, and hybrid cloud networking.
  • Proven experience supporting highly available production environments.
  • Experience with SLOs, incident management, automation, and operational excellence.
  • Proficiency in Python, Bash, or similar scripting languages.
  • Strong communication and collaboration skills.

Preferred Qualifications:
  • Google Cloud Professional Cloud Network Engineer
  • Google Cloud Professional Cloud Architect
  • Google Cloud Professional Cloud DevOps Engineer
  • Certified Kubernetes Administrator (CKA)
  • Experience with GitOps, FinOps, AI/ML workloads, and hybrid cloud platforms.

Pay is competitive and based on a number of job-related factors, including skills and experience. The starting pay rate/range at time of hire for this position in New York is $133,661.00 - $219,586.00 / year. For other locations, please inquire with your recruiter. The rates/ranges provided herein are the anticipated pay at the time of hire, and do not reflect future job opportunity.

We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.

Similar Jobs

More Jobs at Optimum

  • Site Reliability Engineer III, GCP
    $133K — $219K *
    Bethpage, NY 11714 (Nassau County)
    Information Technology
    In-Person
  • Staff Analyst
    $143K — $236K *
    Toronto, ON M3C 0E3
    Business Services
    In-Person
  • Sr. FP&A Analyst
    $77K — $126K *
    New York, NY 10025 (New York County)
    Finance & Insurance
    In-Person
  • Sr. FP&A Analyst
    $77K — $126K *
    Long Island City, NY 11101 (Queens County)
    Finance & Insurance
    In-Person
  • Project Manager
    $64K — $105K *
    Bethpage, NY 11714 (Nassau County)
    Business Services
    In-Person

More Information Technology Jobs

Find similar Site Reliability Engineer III, GCP jobs: