Site Reliability Engineer

Rainforest

• $110K — $130K *
Finance & Insurance
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 3+ years of experience in Site Reliability Engineering (SRE), DevOps, or cloud infrastructure roles, with a preference for startup experience.
  • Hands-on experience with cloud infrastructure, particularly in AWS, Google Cloud, or Azure.
  • Proficient in Infrastructure as Code (IaC) tools such as Terraform or CloudFormation.
  • Production experience with container orchestration, notably Kubernetes or ECS.
  • Experience in building CI/CD pipelines using GitLab or similar tools.
  • Strong understanding of monitoring and observability practices.
  • Proficiency in at least one modern programming language like Python, Java, Go, or Ruby.

Responsibilities

  • Own and scale AWS-based cloud infrastructure using Terraform and IaC.
  • Build and optimize Elastic Kubernetes Service (EKS) and serverless environments for core payments services.
  • Design and maintain CI/CD pipelines with GitLab for efficient deployments.
  • Implement monitoring and observability tools to ensure high uptime and rapid incident resolution.
  • Automate infrastructure and operational processes to enhance delivery speed.
  • Collaborate with application engineers to improve system performance and reliability.
  • Lead incident response efforts and facilitate postmortems for continuous improvement.

Benefits

  • Comprehensive health benefits package.
  • Unlimited paid time off.
  • Paid parental leave.
  • Fun and flexible working environment.
  • Commitment to investing in employee development and culture.
Full Job Description
Who we're looking for

We're looking for a proactive, hands-on Site Reliability Engineer who thrives in building and scaling cloud infrastructure in fast-moving startup environments. You're someone who enjoys owning systems end-to-end - from infrastructure design to production reliability - and partnering closely with engineers to ship secure, scalable payment platforms. You bring a strong technical foundation, a problem-solving mindset, and a passion for automation, performance, and continuous improvement. If you're excited about making a real impact in fintech, and helping shape SRE practices as the company grows, you'll feel right at home at Rainforest.

What are some of the high-impact opportunities you'll tackle?
  • Owning and scaling Rainforest's Amazon Web Services (AWS)-based cloud infrastructure using Terraform and infrastructure-as-code (IaC) orchestration
  • Building, operating, and continuously improving Elastic Kubernetes Service (EKS) and serverless environments that support our core payments services
  • Designing and maintaining modern CI/CD pipelines with GitLab to enable fast, safe deployments
  • Implementing and evolving monitoring, alerting, and observability to ensure high uptime and quick incident resolution using tools like OpenTelemetry, Prometheus, and New Relic
  • Automating infrastructure and operational processes to eliminate manual work and accelerate delivery
  • Working side-by-side with application engineers to improve system performance, reliability, and scalability
  • Leading incident response efforts, conducting postmortems, and driving continuous improvement
  • Helping to define and roll out SRE best practices, including SLIs, SLOs, and error budgets as the company scales
  • Optimizing for cost, security, and compliance in a regulated fintech environment
  • Supporting and scaling Postgres database infrastructure using AWS RDS offerings (Global Aurora)

This opportunity is for you if you have / are:
  • 3+ years of experience in SRE, DevOps, or cloud infrastructure roles (startup or high-growth experience a plus)
  • Passion for building reliable systems that scale with the business
  • Strong hands-on experience with cloud infrastructure (AWS, Google Cloud, Azure)
  • Deep experience with IaC using tools such as Terraform, OpenTofu, Terragrunt, and CloudFormation
  • Solid production experience with container orchestration (Kubernetes, ECS)
  • Experience building CI/CD pipelines using tools like GitLab and GitHub Actions
  • Strong understanding of monitoring and observability principles and design and providing dashboards, visualizations and alerts
  • Proficiency in at least one modern programming language (e.g., Python, Java, Go, or Ruby).
  • Bachelor's degree or equivalent work experience in the areas of Information Science, Computer Science, or related disciplines is preferred

We offer a comprehensive health benefits package, unlimited paid time off, paid parental leave, a fun and flexible working environment, and continuously invest in our people and our culture.

Similar Jobs

More Jobs at Rainforest

More Finance & Insurance Jobs

Find similar Site Reliability Engineer jobs: