One, Inc

Site Reliability Engineering Lead

One, Inc$135K — $160K *
US-AnywhereRemote in United States
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Proficient in writing production-quality code in languages like Python or Go.
  • Experience with reliability for production systems in high-availability environments, preferably fintech.
  • Background in a high-scale engineering organization with a solid engineering culture.
  • Recent experience managing a small engineering team while maintaining technical involvement.
  • Hands-on experience with observability tools and incident management processes.

Responsibilities

  • Lead a senior SRE team, focusing on strategic reliability work rather than reactive responses.
  • Write production code and develop automation and tooling actively alongside the team.
  • Establish service ownership standards, including alerting and runbooks, and conduct health reviews.
  • Manage incident responses, drive remediation efforts, and build self-service tools for incident resolution.
  • Design a sustainable on-call system through staffing and automation, minimizing stress on team members.
  • Collaborate with engineering leaders to enhance reliability standards and maintain clear roles between SRE and service teams.

Benefits

  • Competitive base salary, stock options, and health benefits starting Day 1.
  • 401(k) plan with company matching contributions.
  • Fully remote working option, with flexible time off.
  • Opportunity for professional growth within a mission-driven and inclusive culture.
Full Job Description
The Role

Reliability at OnePay is a shared engineering responsibility, and this role sits at the center of it as a hands-on technical leader. As our Site Reliability Engineering Manager, you will lead a small, senior SRE team and lead it from the front. You will still write code, build automation and tooling yourself, and stay in the rotation when it counts. You will set how your team works, raise the reliability bar for the services you support, and stay close enough to the systems to make sound technical calls and earn your team's trust by doing the work alongside them. You will report to our Head of Platform and Data Engineering and help keep OnePay dependable for the millions of Americans who trust us with their money.

This is a builder's leadership role. You lead a team, but you are measured as much by what you and your team ship as by the operating model you put around it. This is a coding and building role, not a tooling-procurement or pure people-management seat.

What You'll Do
  • Lead a senior SRE team. Bring planning, clarity, and career growth to a small, strong team, and focus them on high-leverage reliability work over manual, ticket-driven response.
  • Stay hands-on. Write production code, and build automation, tooling, and golden paths yourself. Contribute to your team's codebase, do design and code review, and set the technical bar by example.
  • Define service ownership, alerting, dashboards, runbooks, and deploy and rollback safety, and run regular service-health reviews for the systems your team supports.
  • Stay in the incident loop. Run incident command when needed, drive verified remediation, and prevent recurrence. Help build self-service incident tooling and runbook automation so reliability scales without heroics.
  • Build sustainable on-call. Design durable coverage through staffing, handoffs, and automation rather than open-ended volunteer hours.
  • Partner across engineering. Work with service teams and engineering leaders to raise reliability standards, holding clear boundaries between SRE enablement and service-team ownership.
You Bring
  • You still code today. Strong software engineering foundation, and you write production-quality code in a modern language such as Python or Go. You are comfortable being assessed on your coding.
  • Real reliability depth for production systems and services at scale, ideally in a consumer, fintech, or other high-availability environment.
  • You built your foundation at a high-bar, high-scale engineering organization with a strong engineering culture.
  • Recent people leadership. You have managed a small engineering team within roughly the last few years and want to keep leading while staying technical.
  • Genuinely hands-on with observability, alerting, safe deploys, automation, and incident command. You build these things, not just specify them.
  • Calm, credible communication with your team, partner teams, and during live incidents.


What We Offer
  • Competitive base salary, stock options, and health benefits from Day 1
  • 401(k) plan with company match
  • Fully Remote (US), flexible time off (FTO), and opportunities for growth
  • A high-growth, mission-driven, inclusive culture where your work has real impact


Standard Interview Process

Our process varies by role. Most candidates go through:
  • AI-assisted initial screen
  • Interview with Talent Partner
  • Technical or Hiring Manager Interview
  • Team Interview
  • Executive Interview
  • Offer!


About One, Inc

One, Inc. provides software solutions for the insurance industry. The company offers a cloud-based policy administration system that automates the process of selling and servicing insurance policies. The company's platform enables insurance companies to manage policies, claims, billing, and analytics. The company's customers include insurance carriers, managing general agencies, and brokers. The company was founded in 2005 and is headquartered in Santa Monica, California.
Learn more about One, Inc
Size
200 employees
Industry
Founded
2012

Similar Jobs

More Jobs at One, Inc

More Information Technology Jobs

Find similar Site Reliability Engineering Lead jobs: