MeridianLink, Inc.

Senior Site Reliability Engineer

MeridianLink, Inc.$120K — $145K *
US-AnywhereRemote in United States
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree and 4-6 years of related experience or equivalent work history.
  • 5+ years in DevOps, site reliability, or platform operations with a focus on production systems.
  • 3+ years experience specifically with AWS, emphasizing serverless services like Lambda and S3.
  • Strong SQL skills, particularly in PostgreSQL including operations, backups, and query optimization.
  • Proficiency in scripting languages like TypeScript, Python, and bash for automation tasks.
  • Solid understanding of Linux, DNS, TLS, and Docker, with familiarity in GitHub Actions and IaC tools.
  • Experienced in production monitoring, alerting, and incident response methodologies.

Responsibilities

  • Oversee daily administration of AWS services and database management, primarily PostgreSQL.
  • Manage backup strategies across databases and S3, regularly testing recovery plans.
  • Monitor production environments proactively using CloudWatch and alert systems to address issues ahead of user impact.
  • Lead the debugging of production incidents, creating and updating runbooks and participating in on-call duties.
  • Optimize infrastructure for easy deployment and scalability, maintaining accurate infrastructure as code.
  • Cultivate team knowledge by sharing insights on production operations, promoting a culture of growth.

Benefits

  • Flexible work hours and remote work options.
  • Opportunities for professional development and continuing education.
  • Collaborative team environment focused on innovation and best practices.
  • Participation in a culture that encourages knowledge sharing and personal growth.
Full Job Description
As a Senior Site Reliability Engineer on our cloud engineering team, you'll keep our production environment healthy, secure, and running smoothly. This is an operations-focused role: you'll own the day-to-day administration of our AWS accounts and databases, backup posture across our data stores, and production monitoring and debugging for a fully serverless platform. Your work will span the operational side of the software development life cycle - from deployment to maintenance and updates - always striving for continuous improvement. You'll keep our infrastructure clean, easily deployable, and scalable, creating a stable operating environment for the whole team.

Responsibilities
  • Own day-to-day administration across AWS services, accounts, and access, as well as database administration across PostgreSQL and our other data stores.
  • Own backup posture across databases, S3 buckets, and queues; verify restores regularly and maintain a tested disaster recovery plan.
  • Proactively monitor production - CloudWatch dashboards, metric alarms, log-based metrics, and Slack alerting - addressing operational issues before they impact users.
  • Lead production debugging and incident response: build and maintain runbooks, participate in the on-call rotation, and resolve queue and dead-letter-queue failures through retry, redrive, and recovery.
  • Continuously refine our infrastructure to ensure it is easily deployable and scalable: keep infrastructure as code (SST/Pulumi) accurate, retire unused infrastructure, and keep cost visible and justified.
  • Share your knowledge of production operations with the team, fostering a culture of learning and growth.

Qualifications: Knowledge, Skills, & Abilities
  • Bachelor's degree and 4-6 years of related experience or equivalent work experience.
  • 5+ years of experience in DevOps, site reliability, or platform operations, with significant responsibility for production systems.
  • 3+ years of hands-on experience with AWS, with an emphasis on serverless services (Lambda, SQS, EventBridge, CloudWatch, S3).
  • Strong database administration experience: PostgreSQL operations, backup and recovery, and query performance; comfort administering other data stores.
  • Proficiency in scripting languages such as TypeScript, Python, and bash for production automation and operational tooling.
  • Strong understanding of Linux, DNS, TLS, Docker, GitHub Actions, and infrastructure as code (SST, Pulumi, or Terraform).
  • Experience with production monitoring and alerting, incident response, and on-call ownership.

About MeridianLink, Inc.

MeridianLink, Inc. is a technology company that provides software solutions for the financial services industry. The company's products include loan origination, account opening, and credit reporting software, as well as compliance and risk management tools. MeridianLink's software is used by banks, credit unions, and other financial institutions to streamline their operations and improve their customer experience. The company was founded in 1998 and is headquartered in Lake Forest, California.
Learn more about MeridianLink, Inc.
Size
1,000 employees
Market Cap
$1 billion
Industry
NASDAQ

Similar Jobs

More Jobs at MeridianLink, Inc.

More Information Technology Jobs

Find similar Senior Site Reliability Engineer jobs: