Branch

Senior Site Reliability Engineer (SRE)

Branch$175K — $185K *
US-AnywhereRemote in United States
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in engineering or equivalent experience
  • 3+ years in site reliability engineering
  • Hands-on experience with Java / Spring Boot services in production
  • Proficient in Terraform, Go, Java, Gradle, Docker, OpenTelemetry, and Kubernetes

Responsibilities

  • Partner with Developers to enhance service performance through testing and release processes
  • Design and implement infrastructure, monitoring, and standards for applications
  • Support services through design, development, load testing, and deployment phases
  • Monitor and evaluate key performance indicators for system health
  • Establish SLIs, SLOs, and error budgets with service owners and promote their adoption
  • Optimize platform performance, resilience, and efficiency under load
  • Engage in incident response and root cause analysis

Benefits

  • Market-leading medical, dental, and vision insurance
  • Stock options
  • Free Premium-Tier Origin Financial Wellness subscription
  • Monthly home-office stipend
  • 401k plan with TransAmerica
  • 12-weeks paid parental leave for birthing and non-birthing parents
  • Flexible time off + sick and safe time
  • 11 paid company holidays
Full Job Description
About the role:

As the Senior Site Reliability Engineer, you will lead Branch's effort to achieve greater reliability, performance, scalability, capacity and observability of our platform through automation and tooling. You will also participate in and improve the software development and deployment life cycles as well as develop and implement technical best practices.

Responsibilities include, but are not limited to:
  • Partner with Developers to produce high-performing and robust services through rigorous testing and release procedures
  • Design infrastructure, monitoring, processes, and standards for systems and applications
  • Support services through design, development, load testing, and launch phases
  • Develop, measure, and monitor key performance and service level indicators including availability, latency, and overall system health
  • Define and establish SLIs, SLOs, and error budgets with service owners, and drive adoption across platform teams
  • Profile and optimize platform performance, resilience, and efficiency, including latency, throughput, and capacity planning under load
  • Participate in incident response and root cause analysis
  • Remediate tasks and develop preventative and automated measures to meet SLAs/SLOs/SLIs
  • Manage monitoring services utilized by applications

Qualifications (required):
  • Bachelor's degree in an appropriate engineering discipline or equivalent experience required
  • 3+ years experience in site reliability engineering
  • Strong hands-on experience building and operating Java / Spring Boot services in production
  • Experience with Terraform, Go, Java, Gradle, Docker, OpenTelemetry and Kubernetes

Qualifications (preferred):
  • Experience in scripting using Python and Bash
  • Experience developing and maintaining Kubernetes Operators
  • Exposure to Google Pub/Sub, Redis, Prometheus, Grafana, Google Spanner and MySQL
  • Experience with JVM performance tuning and profiling
  • Experience working on GCP platforms

Compensation:

The base salary range for this role is $175-185k. The salary range displayed reflects an average base salary range for the position across all the U.S. The base salary offered to an applicant could be higher or lower based on each applicant's specific skill set, depth of experience, relevant education or training, etc.

Location:

This position is classified as REMOTEwithin the United States of America.

We are unable to hire candidates located outside of the domestic U.S.

Benefits:
  • Market-leading medical, dental, and vision insurance
  • Stock options
  • Free Premium-Tier Origin Financial Wellness subscription
  • Monthly home-office stipend
  • 401k (TransAmerica)
  • 12-weeks paid parental leave for birthing and non-birthing parents
  • Flexible time off + sick and safe time
  • 11 paid company holidays

About Branch

Branch is a mobile-first technology company that provides financial services for hourly workers. The company's mission is to create a world where working Americans can grow financially. Branch offers a mobile app that allows users to access earned wages, budgeting tools, and financial wellness resources. Branch was founded in 2015 by Atif Siddiqi and has raised over $300 million in funding to date.
Learn more about Branch
Size
500 employees
Industry
Founded
2014

Similar Jobs

More Jobs at Branch

More Information Technology Jobs

Find similar Senior Site Reliability Engineer (SRE) jobs: