The RoleReliability at OnePay is a shared engineering responsibility, and this role sits at the center of it as a hands-on technical leader. As our Site Reliability Engineering Manager, you will lead a small, senior SRE team and lead it from the front. You will still write code, build automation and tooling yourself, and stay in the rotation when it counts. You will set how your team works, raise the reliability bar for the services you support, and stay close enough to the systems to make sound technical calls and earn your team's trust by doing the work alongside them. You will report to our Head of Platform and Data Engineering and help keep OnePay dependable for the millions of Americans who trust us with their money.
This is a builder's leadership role. You lead a team, but you are measured as much by what you and your team ship as by the operating model you put around it. This is a coding and building role, not a tooling-procurement or pure people-management seat.
What You'll Do- Lead a senior SRE team. Bring planning, clarity, and career growth to a small, strong team, and focus them on high-leverage reliability work over manual, ticket-driven response.
- Stay hands-on. Write production code, and build automation, tooling, and golden paths yourself. Contribute to your team's codebase, do design and code review, and set the technical bar by example.
- Define service ownership, alerting, dashboards, runbooks, and deploy and rollback safety, and run regular service-health reviews for the systems your team supports.
- Stay in the incident loop. Run incident command when needed, drive verified remediation, and prevent recurrence. Help build self-service incident tooling and runbook automation so reliability scales without heroics.
- Build sustainable on-call. Design durable coverage through staffing, handoffs, and automation rather than open-ended volunteer hours.
- Partner across engineering. Work with service teams and engineering leaders to raise reliability standards, holding clear boundaries between SRE enablement and service-team ownership.
You Bring- You still code today. Strong software engineering foundation, and you write production-quality code in a modern language such as Python or Go. You are comfortable being assessed on your coding.
- Real reliability depth for production systems and services at scale, ideally in a consumer, fintech, or other high-availability environment.
- You built your foundation at a high-bar, high-scale engineering organization with a strong engineering culture.
- Recent people leadership. You have managed a small engineering team within roughly the last few years and want to keep leading while staying technical.
- Genuinely hands-on with observability, alerting, safe deploys, automation, and incident command. You build these things, not just specify them.
- Calm, credible communication with your team, partner teams, and during live incidents.
What We Offer- Competitive base salary, stock options, and health benefits from Day 1
- 401(k) plan with company match
- Fully Remote (US), flexible time off (FTO), and opportunities for growth
- A high-growth, mission-driven, inclusive culture where your work has real impact
Standard Interview ProcessOur process varies by role. Most candidates go through:
- AI-assisted initial screen
- Interview with Talent Partner
- Technical or Hiring Manager Interview
- Team Interview
- Executive Interview
- Offer!