Site Reliability Engineer

Seekr

$120K — $145K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years in Site Reliability Engineering, preferably managing SaaS environments.
  • 5+ years experience with Linux systems and network protocols.
  • Proficient with self-hosted monitoring and logging tools such as ELK, Prometheus, InfluxDB, Grafana.
  • Skilled in load testing applications.
  • Proficient in programming/scripting languages like Python, Ruby, Bash, Java.
  • Adept in container technologies including Docker and Kubernetes.
  • Experienced with automation tools for systems configuration management, focusing on Puppet and Terraform.

Responsibilities

  • Design, architecture, and implementation of systems to meet SLAs.
  • Lead solutions for operational and reliability challenges, identifying system failures.
  • Collaborate with development teams to build and monitor services.
  • Conduct load testing on applications across the organization.
  • Own observability stack management and improvement.
  • Participate in incident response practices and procedures.
  • Automate operational processes to improve efficiency.
  • Troubleshoot live systems and deployment issues.

Benefits

  • Work with a collaborative team on impactful AI challenges.
  • Equity ownership through RSUs for long-term success sharing.
  • Unlimited PTO with 14 paid holidays to encourage work-life balance.
  • Flexible hybrid work environment with offices in Reston, VA, and Austin, TX.
  • Comprehensive total rewards including base salary and bonuses.
  • 401(k) plan with employer matching for retirement savings.
  • Immediate access to health benefits for employees and families.
  • Paid parental leave for new family additions.
Full Job Description
We are looking for a talentedSenior Site Reliability Engineer to join our team to deliver world class search technologies to mobile devices. You will be working with a smart team of Engineers to lead and drive the stability, reliability, and observability of all Seekr's software, powering Seekr's powerful search technology.

From your first day, you will make a valuable - and valued - contribution. We are a fast-growing company where no one is a bystander. We offer you the opportunity to delight millions of consumers around the world while gaining meaningful experience across a variety of disciplines.

Duties and Responsibilities
  • Participate in the design, architecture and implementation of systems, software, networks, and services required to keep the Seekr platform within SLAs.
  • Lead development of solutions to complex operational and reliability challenges and proactive detection of system failures and scalability issues in production.
  • Work in close collaboration with software development teams to build, monitor and triage our services.
  • Load testing applications across the organization.
  • Take ownership of our observability stack.
  • Share ownership of our incidence response practices and procedures.
  • Identify and remedy operational inefficiencies through automation.
  • Troubleshoot and assist with incident response on live systems and deployment related issues.

Qualifications and Skills Required
  • 5+ years working in Site Reliability Engineering preferably managing SaaS environments.
  • 5+ years of experience working with Linux systems, familiar with Linux fundamentals including network protocols.
  • Experienced with self-hosted monitoring, metrics, and centralized logging tools (ELK, Prometheus, InfluxDB, Grafana; required).
  • Skilled in load testing applications.
  • Proficient in programming and/or scripting languages (Python, Ruby, Bash, Java).
  • Adept with container technologies Docker and Kubernetes (required).
  • Demonstrated ability in systems configuration management using automation tools such as Puppet, Chef, Ansible, and Terraform (Puppet and Terraform preferred).
  • Knowledgeable in monitoring Kubernetes, Elasticsearch, Kafka, and Aerospike (preferred).
  • Versed in Git (on-prem or GitHub); familiarity with GitLab and ArgoCD is a bonus.
  • Comfortable working within a hybrid cloud environment, with on-prem experience preferred.
  • Proven track record of delivering results on time and with high quality.
  • Reliable in maintaining SLAs.
  • Effective communicator with the ability to work across multiple business and technical teams.

Qualifications and Skills Desired
  • Advanced degree is nice to have

#LI-CT1

Company Benefits:
  • Meaningful Mission & Impact - Work with a deeply talented, collaborative team solving some of the toughest AI challenges that matter.
  • Equity Ownership - RSUs that let you share directly in Seekr's long-term success and growth.
  • Time Off That Respects Real Life - Unlimited PTO plus 14 paid company holidays to truly recharge.
  • Work Your Way - A flexible hybrid work environment with offices in Reston, VA and Austin, TX.
  • Competitive Total Rewards - A role-appropriate compensation structure that supports long-term growth, including base salary, bonuses, or commission plans depending on role.
  • 401(k) with Company Match - Build your future with a retirement plan that includes employer matching.
  • Comprehensive Health & Wellness - Medical, dental, vision, and life insurance coverage starting day one-for you and your family.
  • Parental Leave - Paid parental leave to support employees as they welcome a new child through birth, adoption, or foster placement.


Similar Jobs

More Jobs at Seekr

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: