10+ years of experience in distributed and cloud-native software development at a large-scale service
Deep understanding of scalable, resilient distributed systems and cloud-native architectures
Proven ability to influence and lead large projects across multiple teams
Strong collaborative skills with diverse stakeholders and subject matter experts
High level of autonomy and a proactive, entrepreneurial mindset
Experience as a technical lead and mentor across engineering teams
Responsibilities
Collaborate with leaders and engineers to identify opportunities for improving reliability
Scope, design, implement, and deploy robust distributed services with a focus on tradeoffs
Innovate and implement new products and prototypes to enhance service resiliency
Mentor and develop the next generation of technical leaders at Stripe
Contribute to engineering strategy, tooling, processes, and culture
Uphold high engineering standards and improve codebase and processes
Benefits
Remote work flexibility within the U.S. and Canada
Opportunity to work on cutting-edge, greenfield solutions
Engagement in high-impact projects that shape the future of Stripe's architecture
Collaborative environment with diverse teams and stakeholders
Focus on personal and professional growth through mentorship opportunities
Full Job Description
About the team
In this role, you will be joining the High Availability and Disaster Recovery team. At Stripe, availability is a core feature of our products. This team designs and builds new solutions to allow latency-critical, stateful applications to survive any type of disaster. We build distributed systems on top of unreliable architecture to provide highly available and resilient customer solutions. This team is creating greenfield solutions which will serve as the basis for Stripe's architecture 5, 10, or 20 years into the future.
This is a distributed team with many remote engineers. You are encouraged to apply if you meet the minimum requirements and are able to work from anywhere in the United States or Canada. What you'll do
You will help develop our global architecture by combining less-available components and data centers into a highly available and resilient whole. You will work on latency-critical solutions where every millisecond matters and data redundancy is a hard requirement. You will learn quickly and work on a broad range of problems - one day may be investigating Mongo write concerns, the next may be minimizing cross-region TLS handshakes, followed by developing new systems to automate disaster detection and failovers. Your work will enable Stripe to increase the GDP of the internet by providing uptime and data protection which have historically been impossible. Responsibilities
Actively work with leaders and engineers across the company to understand and identify opportunities to improve Stripe's reliability posture
Scope, design, implement, and deploy robust distributed services, making appropriate tradeoffs between reliability, throughput, latency, resiliency, engineering velocity and cost
Innovate, design and implement new products and prototypes to improve service resiliency, engineering velocity and management at scale
Mentor and grow the next generation of technical leaders at Stripe
Contribute to engineering strategy, tooling, processes, and culture
Uphold our high engineering standards and improve our codebase and processes
Who you are
We're looking for someone who meets the minimum requirements to be considered for the role. If you meet these requirements, you are encouraged to apply. The preferred qualifications are a bonus, not a requirement. Minimum requirements
Have 10 years or more of distributed and cloud native software development experience at a highly-scaled or hyperscaler service
Have a deep understanding of how to build scalable, resilient, distributed systems, cloud native architectures, and mission critical systems.
Have experience influencing, planning, scoping, and leading large projects across many teams
Thrive in a collaborative environment involving diverse stakeholders and subject matter experts
Have a high level of autonomy and responsibility, and think of yourself as entrepreneurial, proactive and self-driven
Have a strong history of being a technical lead and mentor across several engineering teams
About Stripe
Stripe is a technology company that builds economic infrastructure for the internet. Businesses of every size—from new startups to public companies—use our software to accept payments and manage their businesses online. Stripe helps new companies get started and grow their revenues, and established businesses accelerate into new markets and launch new business models. Stripe powers businesses all over the world, from the new startup that just launched yesterday to the Fortune 500 companies that we all know and love. Stripe is headquartered in San Francisco, with offices in Dublin, London, Paris, Singapore, Tokyo, and more.