Senior Site Reliability Engineer

Ivo

$150K — $180K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Minimum 5 years of professional experience in infrastructure engineering
  • Ability to think in failure modes
  • Proven experience in designing resilient systems
  • Ability to translate contractual clauses into infrastructure constraints
  • Strong sense of urgency and resourcefulness

Responsibilities

  • Own uptime, reliability, and performance end-to-end
  • Define and enforce SLI, SLO, and SLA targets
  • Design effective failover and disaster recovery systems
  • Implement data residency requirements like geo-fencing
  • Ensure security controls that pass audits and maintain performance
  • Build observability systems to analyze failures
  • Lead incident response and write engaging postmortems

Benefits

  • Competitive compensation package based on experience
  • Equity options for ownership in a fast-scaling company
  • Relocation assistance and visa support for successful candidates
  • Comprehensive medical, dental, and vision health plans
  • Flexible spending accounts and life insurance coverage
  • 401(k) program for future savings
  • Commuter benefits for easier travel to the office
  • Unlimited PTO for work-life balance
  • Office perks including catered lunch, snacks, gym access, and a dog-friendly environment
Full Job Description
The Role: Why, What and the Who

Why? Infrastructure Engineers build the foundation for Ivo's entire platform.

Customers are cagey about their contracts, so each customer gets their isolated environment with containers, database, VPC, etc. Things break. Regions go down. Cloud and LLM providers have "incidents." Customers still expect us to hit our SLAs.

What? We're looking for an Senior Site level Reliability Engineer as part of Infrastructure team to:
  • Own uptime, reliability, and performance end-to-end
  • Define and enforce SLI, SLO and SLA targets (and make sure we don't get paged in the wee hours)
  • Design failover + disaster recovery that actually works in real scenarios
  • Turn data residency requirements into real systems (geo-fencing, regional isolation, etc.)
  • Implement security controls that pass audits and don't slow the product to a crawl
  • Build observability that answers: what, why and how often it broke ?
  • Lead incident response + write postmortems that make people actually read


Who? We need someone who:
  • Minimum 5 years of professional experience
  • Thinks in failure modes
  • Designs systems that keep working anyway
  • Can translate "this clause in a contract" into actual infrastructure constraints

This isn't a "keep the lights on" role. You'll be building the system that keeps the company running. In addition to helping us run a solid, high-performance distributed system, we'd love someone who's as excited about LLMs as we are. You'd be deeply embedded into the engineering team and highly encouraged to push the frontiers.

Ivo might be a good fit for you if you:
  • You love writing code, but you love having impact more: We're a team of engineers at heart, but our #1 goal is building the best possible product. That means making pragmatic choices and looking for 80/20 solutions.
  • Would describe yourself as being relentlessly resourceful.
  • You have a strong internal sense of urgency. You have a bias towards doing things *today*, rather than tomorrow.
  • Experience working in a startup environment is preferred but not required.
  • Are excited about the adventure of building a company!


What We Offer
  • Competitive Compensation: Final offer details are determined based on experience, expertise, and overall fit.
  • Equity: Meaningful ownership in a company that's scaling fast
  • Relocation and Visa Support: We also offer relocation assistance for successful applicants moving to SF, as well as support for visa and green card applications where applicable.
  • Health & Wellness: Comprehensive medical, dental, and vision plans to suit the needs of you and your family.
  • Flexible Spending & Insurance: Access to HSA and FSA accounts, plus life insurance coverage.
  • 401(k) Program: Save for the future with our 401(k) program.
  • Commuter Benefits: We help make getting to and from the office easier and more convenient.
  • Unlimited PTO: So you can take the time you need to recharge, stay healthy, and bring your best self to work.
  • Office Perks: Enjoy a vibrant Downtown San Francisco office with catered lunch five days a week, premium snacks and coffee, an in-building gym, and a dog-friendly environment.


FAQ
  • What stage of growth is Ivo at?: We launched in early access in 2023. Since then, we've had an incredible response from the market and are growing rapidly. We 6x'd in ARR in the last 12 months. Our clients include companies like Uber, Reddit, IBM, Canva, Pinterest, WordPress, and more. We're happy to share more details with candidates who go through our interview process.
  • Is this a chill gig?: Startups are very hard, especially if they're growing fast. You'll have a ton of responsibility, and there's always an enormous amount of stuff to do. It's hard work but the payoff is uncapped.
  • Can I work remotely?: We require candidates to work with us in-person 5 days a week in our San Francisco office.

Similar Jobs

More Jobs at Ivo

More Information Technology Jobs

Find similar Senior Site Reliability Engineer jobs: