Site Rel Eng III, GCP

Optimum

$133K — $219K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years in SRE, Platform Engineering, Cloud Infrastructure, DevOps, or Networking
  • 5+ years with production workloads on Google Cloud Platform
  • Expertise in GCP, GKE, Kubernetes, and Terraform
  • Strong knowledge of cloud networking and IAM practices
  • Experience with incident management and operational excellence
  • Proficiency in scripting languages like Python or Bash
  • Excellent teamwork and communication skills

Responsibilities

  • Design and operate enterprise-scale GCP infrastructure
  • Build solutions with GKE, Cloud Run, and other GCP services
  • Architect and support advanced cloud networking solutions
  • Develop Infrastructure as Code and automate workflows with CI/CD
  • Define SLOs and enhance observability through automation
  • Lead incident response and reliability improvement efforts
  • Implement security and governance best practices across cloud environments
  • Mentor engineering teams and establish SRE standards

Benefits

  • Opportunities for professional development and certification
  • Collaborative and innovative work environment
  • Focus on engineering excellence and automation
  • Access to advanced technologies and cloud solutions
  • Support for career advancement within the organization
Full Job Description
Job Summary

Optimum is seeking a Lead Site Reliability Engineer (SRE) to design, build, and operate secure, scalable, and highly available platforms on Google Cloud Platform (GCP). This hands-on engineering role combines cloud infrastructure, networking, automation, Kubernetes, and reliability engineering to support critical enterprise and customer-facing applications.

The ideal candidate brings deep expertise in GCP, cloud networking, Infrastructure as Code, observability, and incident management, along with a passion for improving reliability through automation and engineering excellence. You are a hands-on engineer with deep GCP and cloud networking expertise who enjoys solving complex operational challenges through automation, platform engineering, and reliability-focused design. You combine technical depth with leadership to build resilient, scalable cloud platforms that enable teams to deliver software with confidence.

Responsibilities

  • Design and operate enterprise-scale GCP infrastructure and platforms.
  • Build and manage solutions using GKE, Cloud Run, Compute Engine, Cloud SQL, Pub/Sub, Cloud Storage, and Cloud Monitoring.
  • Architect and support cloud networking, including Network Connectivity Center (NCC), Interconnect, Cloud VPN, Cloud Router, Shared VPC, and Private Service Connect.
  • Develop Infrastructure as Code using Terraform and automate deployment and operational workflows through CI/CD pipelines.
  • Define and manage SLOs, improve observability, and reduce operational toil through automation.
  • Lead incident response, root cause analysis, and reliability improvement initiatives.
  • Implement cloud security, IAM, governance, and operational best practices.
  • Partner with application, network, and security teams to deliver reliable cloud-native platforms.
  • Mentor engineers and help establish platform and SRE standards across the organization.


Qualifications

  • 8+ years of experience in SRE, Platform Engineering, Cloud Infrastructure, DevOps, or Network Engineering.
  • 5+ years of hands-on experience operating production workloads on Google Cloud Platform.
  • Strong expertise in:
    • GCP and GKE
    • Kubernetes
    • Terraform
    • Cloud Networking
    • IAM
    • Monitoring and Observability
    • Infrastructure as Code
  • Experience with NCC, Interconnect, Cloud VPN, Cloud Router, Shared VPC, and hybrid cloud networking.
  • Proven experience supporting highly available production environments.
  • Experience with SLOs, incident management, automation, and operational excellence.
  • Proficiency in Python, Bash, or similar scripting languages.
  • Strong communication and collaboration skills.

Preferred Qualifications:
  • Google Cloud Professional Cloud Network Engineer
  • Google Cloud Professional Cloud Architect
  • Google Cloud Professional Cloud DevOps Engineer
  • Certified Kubernetes Administrator (CKA)
  • Experience with GitOps, FinOps, AI/ML workloads, and hybrid cloud platforms.

Pay is competitive and based on a number of job-related factors, including skills and experience. The starting pay rate/range at time of hire for this position in New York is $133,661.00 - $219,586.00 / year. For other locations, please inquire with your recruiter. The rates/ranges provided herein are the anticipated pay at the time of hire, and do not reflect future job opportunity.

We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.

Similar Jobs

More Jobs at Optimum

More Information Technology Jobs

Find similar Site Rel Eng III, GCP jobs: