Site Reliability Engineer

Empower$87K — $123K *
US-AnywhereRemote in United States
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Experience maintaining high availability and resiliency within AWS infrastructure components, including EKS, EC2, RDS, S3, VPC, and others
  • Proficiency with Infrastructure as Code frameworks such as Terraform and CloudFormation
  • Demonstrated experience with containerization and orchestration technologies such as Docker and Kubernetes
  • Experience identifying and mitigating production incidents
  • Strong problem-solving abilities and eagerness to learn

Responsibilities

  • Establish key indicators (SLIs) to measure service performance and build proactive monitoring and alerts
  • Support diverse projects across multiple disciplines and teams
  • Conduct in-depth analysis of issues to identify findings and root causes
  • Ensure high availability, resilience, and scalability of containerized applications in production
  • Document critical systems and create runbooks for incidents
  • Lead capacity planning and right-sizing initiatives
  • Optimize Infrastructure as Code (IaC) and troubleshoot complex system issues

Benefits

  • Medical, dental, vision, and life insurance
  • 401(k) plan with generous matching contributions and financial advisory services
  • Tuition reimbursement up to $5,250/year
  • Generous paid time off, including paid holidays and floating days
  • Paid volunteer time - 16 hours per year
  • Leave of absence programs, including paid parental leave and disability
  • Business Resource Groups (BRGs) for inclusion and collaboration
Full Job Description
What you will do:
  • Establish key indicators (SLIs) measuring the performance of services and build proactive monitor and alerts
  • Support projects of varying complexity and impact across multiple disciplines and teams
  • Conduct in-depth analysis of problems to identify relevant findings and root causes
  • Play a key role in ensuring the high availability, resilience, and scalability of containerized applications in production
  • Document critical systems and create runbooks for incidents
  • Lead capacity planning and right-sizing exercises
  • Maintain and optimize infrastructure as code (IaC)
  • Troubleshoot and resolve complex system and deployment issues
  • Manage observability withinKubernetes, specifically EKS
  • Collaborate with development teams to support releases and create highly scalable, resilient, and maintainable services
  • Work in a GitOps driven environment


What you will bring:
  • Experience maintaining high availability and resiliency within AWS infrastructure components, including EKS, EC2, RDS, S3, VPC, and others
  • Proficiency with Infrastructure as Code frameworks such as Terraform and Cloudformation
  • Demonstrated experience with containerization and orchestration technologies such as Docker and Kubernetes
  • Experience with technologies, systems, networks, and potential gaps that can impact an organization's ability to effectively detect and respond to production incidents
  • Strong problem-solving abilities and a desire to learn


What will set you apart:
  • Bachelor's degree in Computer Science, Information Systems or equivalent experience
  • AWS, Kubernetes, or relevant certifications
  • Experience with observability suites and APM tooling such as DataDog, AppDynamics, New Relic, etc.
  • Strong programming skills in one or more languages such as shell, Go, Python, etc.
  • Experience supporting Java Spring Boot applications
  • Production experience in Kubernetes, especially EKS


This job description is not intended to be an exhaustive list of all duties, responsibilities and qualifications of the job. The employer has the right to revise this job description at any time. You will be evaluated in part based on your performance of the responsibilities and/or tasks listed in this job description. You may be required perform other duties that are not included on this job description. The job description is not a contract for employment, and either you or the employer may terminate employment at any time, for any reason.

What we offer you

We offer an array of diverse and inclusive benefits regardless of where you are in your career. We believe that providing our employees with the means to lead healthy balanced lives results in the best possible work performance.
  • Medical, dental, vision and life insurance
  • Retirement savings - 401(k) plan with generous company matching contributions (up to 6%), financial advisory services, potential company discretionary contribution, and a broad investment lineup
  • Tuition reimbursement up to $5,250/year
  • Business-casual environment that includes the option to wear jeans
  • Generous paid time off upon hire - including a paid time off program plus ten paid company holidays and three floating holidays each calendar year
  • Paid volunteer time - 16 hours per calendar year
  • Leave of absence programs - including paid parental leave, paid short- and long-term disability, and Family and Medical Leave (FMLA)
  • Business Resource Groups (BRGs) - BRGs facilitate inclusion and collaboration across our business internally and throughout the communities where we live, work and play. BRGs are open to all.


Base Salary Range
$87,400.00 - $123,400.00

The salary range above shows the typical minimum to maximum base salary range for this position in the location listed. Non-sales positions have the opportunity to participate in a bonus program. Sales positions are eligible for sales incentives, and in some instances a bonus plan, whereby total compensation may far exceed base salary depending on individual performance. Actual compensation offered may vary from posted hiring range based upon geographic location, work experience, education, licensure requirements and/or skill level and will be finalized at the time of offer.

***For remote and hybrid positions you will be required to provide reliable high-speed internet with a wired connection as well as a place in your home to work with limited disruption. You must have reliable connectivity from an internet service provider that is fiber, cable or DSL internet. Other necessary computer equipment, will be provided. You may be required to work in the office if you do not have an adequate home work environment and the required internet connection.***

Job Posting End Date at 12:01 am on:
08-24-2026

Want the latest money news and views shaping how we live, work and play? Stay in the know with The Currency and sign up for Empower's free newsletter.

About Empower

Empower is a retirement plan recordkeeping financial holding company based in Greenwood Village, Colorado, United States. It is the second-largest retirement plan provider in the United States.
Learn more about Empower

Similar Jobs

More Jobs at Empower

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: