Site Reliability Engineer

Seekr

$120K — $145K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years in Site Reliability Engineering, preferably in SaaS
  • 5+ years of experience with Linux systems and network protocols
  • Proficient with monitoring and logging tools (ELK, Prometheus, Grafana)
  • Skilled in load testing applications
  • Experienced in programming/scripting (Python, Ruby, Bash, Java)
  • Adept with Docker and Kubernetes
  • Knowledge in automation tools (Puppet, Chef, Ansible, Terraform)

Responsibilities

  • Participate in the design and implementation of systems to meet SLAs
  • Lead solutions for operational and reliability challenges
  • Collaborate with development teams for service monitoring and triage
  • Conduct load testing for applications across the organization
  • Take ownership of the observability stack
  • Share responsibilities in incident response practices
  • Identify operational inefficiencies and automate remedies

Benefits

  • Meaningful mission and impact through collaboration on AI challenges
  • Equity ownership via RSUs for long-term growth
  • Unlimited PTO coupled with 14 paid company holidays
  • Flexible hybrid work environment with options in VA and TX
  • Competitive total rewards including various compensation structures
  • 401(k) with employer matching for retirement savings
  • Comprehensive health insurance starting on day one
  • Paid parental leave for new parents
Full Job Description
We are looking for a talentedSenior Site Reliability Engineer to join our team to deliver world class search technologies to mobile devices. You will be working with a smart team of Engineers to lead and drive the stability, reliability, and observability of all Seekr's software, powering Seekr's powerful search technology.

From your first day, you will make a valuable - and valued - contribution. We are a fast-growing company where no one is a bystander. We offer you the opportunity to delight millions of consumers around the world while gaining meaningful experience across a variety of disciplines.

Duties and Responsibilities
  • Participate in the design, architecture and implementation of systems, software, networks, and services required to keep the Seekr platform within SLAs.
  • Lead development of solutions to complex operational and reliability challenges and proactive detection of system failures and scalability issues in production.
  • Work in close collaboration with software development teams to build, monitor and triage our services.
  • Load testing applications across the organization.
  • Take ownership of our observability stack.
  • Share ownership of our incidence response practices and procedures.
  • Identify and remedy operational inefficiencies through automation.
  • Troubleshoot and assist with incident response on live systems and deployment related issues.

Qualifications and Skills Required
  • 5+ years working in Site Reliability Engineering preferably managing SaaS environments.
  • 5+ years of experience working with Linux systems, familiar with Linux fundamentals including network protocols.
  • Experienced with self-hosted monitoring, metrics, and centralized logging tools (ELK, Prometheus, InfluxDB, Grafana; required).
  • Skilled in load testing applications.
  • Proficient in programming and/or scripting languages (Python, Ruby, Bash, Java).
  • Adept with container technologies Docker and Kubernetes (required).
  • Demonstrated ability in systems configuration management using automation tools such as Puppet, Chef, Ansible, and Terraform (Puppet and Terraform preferred).
  • Knowledgeable in monitoring Kubernetes, Elasticsearch, Kafka, and Aerospike (preferred).
  • Versed in Git (on-prem or GitHub); familiarity with GitLab and ArgoCD is a bonus.
  • Comfortable working within a hybrid cloud environment, with on-prem experience preferred.
  • Proven track record of delivering results on time and with high quality.
  • Reliable in maintaining SLAs.
  • Effective communicator with the ability to work across multiple business and technical teams.

Qualifications and Skills Desired
  • Advanced degree is nice to have

#LI-CT1

Company Benefits:
  • Meaningful Mission & Impact - Work with a deeply talented, collaborative team solving some of the toughest AI challenges that matter.
  • Equity Ownership - RSUs that let you share directly in Seekr's long-term success and growth.
  • Time Off That Respects Real Life - Unlimited PTO plus 14 paid company holidays to truly recharge.
  • Work Your Way - A flexible hybrid work environment with offices in Reston, VA and Austin, TX.
  • Competitive Total Rewards - A role-appropriate compensation structure that supports long-term growth, including base salary, bonuses, or commission plans depending on role.
  • 401(k) with Company Match - Build your future with a retirement plan that includes employer matching.
  • Comprehensive Health & Wellness - Medical, dental, vision, and life insurance coverage starting day one-for you and your family.
  • Parental Leave - Paid parental leave to support employees as they welcome a new child through birth, adoption, or foster placement.

Similar Jobs

More Jobs at Seekr

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: