DriveWealth

Manager, Site Reliability Engineering

DriveWealth • $150K — $170K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of experience in managing or leading SRE/DevOps engineers, preferably in fintech or regulated sectors.
  • Strong understanding of Google’s SRE practices, including SLOs, error budgets, and toil reduction methodologies.
  • Expertise in Linux administration, networking principles, and troubleshooting skills.
  • Experience managing production-grade Kubernetes clusters and understanding advanced orchestration patterns.
  • Strong knowledge of AWS services, security implications, and high availability design patterns.
  • Proficiency in scripting and coding, particularly in Python, Golang, Bash, and Ansible.
  • Familiarity with data orchestration tools like Rundeck and Airflow, alongside modern CI/CD practices.

Responsibilities

  • Manage and mentor a team of SRE Automation Engineers, focusing on their career development and technical direction.
  • Lead the development of tooling and automation to reduce manual toil and enhance workflow efficiency.
  • Implement Google’s SRE principles in a regulated environment to enhance reliability and performance.
  • Set standards for Infrastructure as Code using Terraform and oversee GitOps operations via ArgoCD.
  • Review and optimize software architecture and Kubernetes operations for cost and performance across AWS.
  • Drive incident response efforts and promote a blameless culture for post-mortem analysis.
  • Collaborate with engineering leaders to align SRE efforts with overarching business objectives and improve best practices.

Benefits

  • Medical, dental, and vision insurance coverage.
  • 401(k) matching plan to support your retirement savings.
  • Paid parental leave and generous PTO policies for work-life balance.
  • Wellness reimbursement and personal development allowance.
  • Company-provided mobile phone for work efficiency.
Full Job Description
About The Role

As the Manager of Site Reliability Engineering, you'll lead a team of SRE Automation Engineers while remaining a hands-on technical authority for our Brokerage-as-a-Service platform. This isn't a purely people-management seat, you're expected to bring the same principal-level SRE depth to automation design and engineering as an individual contributor, while also building the team, setting technical direction, and developing your engineers' careers.

This role is centered on reducing manual toil through engineering, applying Google's SRE principles: SLOs, error budgets, blameless postmortems, and systematic toil reduction, adapted to a regulated brokerage environment. You'll carry two responsibilities at once: driving the automation agenda, building and orchestrating workflows in Rundeck and Airflow to eliminate repetitive work, and growing your team of SRE Automation Engineers into a high-functioning automation practice. You'll guide the design of internal SRE platforms, automate complex workflows, and ensure our Kubernetes-based and colo ecosystems can handle the demands of global financial markets, while owning the people side of the team: mentorship, performance, and growth, and the day-to-day management of the team's Jira board. While this role includes participation in on-call rotations supporting our 24/7 global operations, your primary mission is to build systems that make manual intervention obsolete, and a team capable of sustaining that mission.

What You'll Do
  • Team Leadership & Development: Manage, mentor, and grow a team of SRE Automation Engineers-setting technical direction, running 1:1s, owning performance management and career development, and managing the team's Jira board to prioritize and track sprint work.
  • Engineering & Automation: Lead the design and development of internal tooling and automation-including Rundeck and Airflow-based orchestration-to eliminate repetitive manual toil and improve developer velocity, staying hands-on with the most complex, highest-leverage automation work yourself.
  • SRE Practice & Governance: Adapt Google's SRE principles to our environment-defining SLIs, SLOs, and error budgets, and using them to guide engineering and operational priorities.
  • Infrastructure as Code: Set architectural standards for modular, reusable IaC using Terraform and oversee GitOps workflows via ArgoCD.
  • Platform Governance: Review software architecture and Kubernetes metrics to ensure high availability, capacity planning, and cost-optimization across AWS regions, and hold the team accountable to those standards.
  • Incident Engineering: Lead incident response for critical events, drive complex root-cause analysis (RCA), and champion a blameless post-mortem culture across the organization.
  • Collaboration & Stakeholder Management: Partner with engineering leadership to align SRE priorities with business goals, and foster adoption of new tools, security standards, and reliability best practices across teams.
You Bring
  • People Leadership: Prior experience managing or leading SRE/DevOps engineers, ideally in a fintech or highly regulated environment. Able to flex between hands-on principal-level engineering and coaching and developing a team.
  • Google SRE Fundamentals: Working knowledge of Google's SRE practices-SLIs/SLOs, error budgets, toil reduction, and blameless postmortems-and experience adapting them to a regulated environment.
  • Linux & Networking Mastery: Proficient in Linux administration with a deep understanding of the TCP/IP stack, OSI model, DNS, and network troubleshooting.
  • FinTech Background: Experience working in highly regulated financial environments or with FIX/API connectivity.
  • Production Kubernetes: Hands-on experience managing production-grade clusters, including RBAC, autoscaling, Helm, and multi-cluster patterns.
  • Cloud Native Expertise (AWS): Strong grasp of AWS core services, security, and high-availability patterns. Proficiency with boto3 and AWS CLI for automation.
  • Modern CI/CD & GitOps: Experience building secure, automated delivery pipelines and operating GitOps workflows (ArgoCD).
  • Code Proficiency: Strong scripting and development skills in Python or Golang, along with Bash and Ansible.
  • Observability: Experience with Grafana/Similar tools, Prometheus, Understanding of logs shipping, management and metric first alerting.
  • Security Mindset: Experience with secrets management, vulnerability scanning, and securing the software supply chain.
  • AI & Prompt Engineering: Familiarity with using LLMs, Public MCPs, or Bedrock Agent Core to enhance SRE workflows.
  • Data & Middleware & Orchestration: Hands-on experience with Rundeck and Airflow for job orchestration and automation, plus experience managing Kafka, MQ, or SQS.
Location

This role is open to candidates in the following locations: Chicago, IL - Hybrid
  • This role is expected to come into the office on a cadence set by the Hiring Manager/Team.
  • If you're not based in the location listed above, this role is not a fit, and we cannot accommodate remote work outside these locations.
  • Applicants must be authorized to work for any employer in the U.S. DriveWealth does not sponsor or take over sponsorship of an employment visa at this time.

Pay Range: $150,000 - $170,000 USD

Working at DriveWealth

We do our best work when we're in the same room. To maintain the speed our partners expect, our New York, Chicago, and Lithuania teams work in office on a hybrid schedule. We've found that being physically side-by-side is the only way to solve complex problems in real-time and stay truly accountable to the products we ship. When you're here, you're working directly with the people making the decisions.

To support that work, we provide competitive compensation, equity, and a 401(k) match. We also offer Medical insurance, Dental insurance, Vision insurance, Disability insurance, and Paid Parental Leave, along with a wellness reimbursement, a company-provided phone, and a personal development allowance. Finally, we value the time you spend away from the office with generous Paid Time Off (PTO) and observed holidays.

How We Think About AI

We leverage AI to work smarter and move faster. We seek AI-curious talent who are proactive about using emerging tools to increase signal quality, reduce friction, and improve outcomes to deliver products faster, provide better service to our partners, and to streamline processes. Your ability to leverage our internal tools and technology to drive results is as important to us as your core domain expertise.

Compensation

Pay is generally based on the level, complexity, responsibility, location, and job duties/requirements of the specific position. We then source candidates with the requisite skills, expertise, education, training, and experience. If you are selected for an interview, please feel welcome to speak to a recruiter about our compensation philosophy and other available benefits. This role is eligible for base, bonus, equity, 401(k) match, and heavily subsidized benefits and perks.

Agency Disclaimer

DriveWealth does not accept agency resumes. Do not forward resumes to our jobs alias, employees, or any other organization location. DriveWealth is not responsible for any fees related to unsolicited resumes.

About DriveWealth

DriveWealth is a financial services company that provides a digital trading platform for retail investors. The company's platform enables investors to buy and sell securities in real-time using fractional shares, making it easier and more affordable for individuals to invest in the stock market. DriveWealth's platform is available in over 150 countries and supports trading in US equities, exchange-traded funds (ETFs), and American depositary receipts (ADRs). The company was founded in 2012 and is headquartered in Chatham, New Jersey.
Learn more about DriveWealth
Size
100 employees
Industry
Founded
2012

Similar Jobs

More Jobs at DriveWealth

  • DriveWealth
    Senior Quantitative Developer
    $220K — $240K *
    Chicago, IL 60629 (Cook County)
    Finance & Insurance
    In-Person
  • DriveWealth
    Brokerage Clearing Specialist
    $100K — $120K *
    Chicago, IL 60629 (Cook County)
    Finance & Insurance
    In-Person
  • DriveWealth
    Delivery Manager
    $175K — $190K *
    Chicago, IL 60629 (Cook County)
    Finance & Insurance
    In-Person
  • DriveWealth
    Delivery Manager
    $175K — $190K *
    Dallas, TX 75217 (Dallas County)
    Finance & Insurance
    In-Person
  • DriveWealth
    Delivery Manager
    $175K — $190K *
    Denver, CO 80219 (Denver County)
    Finance & Insurance
    In-Person

More Information Technology Jobs

Find similar Manager, Site Reliability Engineering jobs: