Forge Global

Manager, Site Reliability Engineer

Forge Global$150K — $220K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of experience leading a Site Reliability Engineering, DevOps, or similar function.
  • 10+ years of overall software engineering or operations experience.
  • Bachelor's degree in Computer Science, Engineering, or related field, or equivalent experience.
  • Experience with large-scale cloud infrastructure and distributed systems.
  • Hands-on experience in observability, monitoring, and incident response.
  • Familiarity with CI/CD processes and operational tooling.
  • Strong technical judgment and communication skills.

Responsibilities

  • Manage the Site Reliability Engineering team for system availability.
  • Drive incident management practices across engineering teams.
  • Build and manage observability infrastructure with the Platform Engineering team.
  • Enhance monitoring coverage and alert quality for quicker response times.
  • Champion reliability best practices throughout engineering.
  • Contribute to technical design and team delivery practices.
  • Collaborate with engineering teams to troubleshoot and improve system reliability.

Benefits

  • Professional development and career growth opportunities.
  • Supportive and collaborative team culture.
  • Hands-on leadership experience in a technical environment.
Full Job Description
The Role:

As an engineering organization, we pride ourselves on engineering as a creative activity. Engineering managers enable engineers to do their best work by maintaining a culture and environment where engineers can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge's SRE team responsible for keeping Forge systems highly available for customers, while partnering closely with Platform, Engineering, Security, Compliance, and Product teams to improve reliability, observability, incident response, and operational maturity. This is an opportunity for a hands-on technical leader who can coach engineers, improve production operations, and help Forge build and run secure, scalable, and highly reliable products.

Responsibilities:
  • Manage Forge's Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.
  • Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.
  • Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.
  • Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.
  • Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.
  • Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.
  • Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.
    • Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.
    • Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.

Qualifications:
  • 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.
  • 10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.
  • Bachelor's degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.
  • Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.
  • Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.
  • Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
  • Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.

Preferred Qualifications:
  • Experience in FinTech, financial services, or another regulated industry.
  • Experience with AWS and/or Azure cloud platforms.
  • Familiarity with Kubernetes, container platforms, infrastructure-as-code, Terraform, Ansible, or similar automation tooling.
  • Experience with observability platforms such as Datadog, CloudWatch, or similar tools.
  • Experience improving developer experience through paved-road platforms, standardization, and self-service infrastructure capabilities.
  • Experience supporting growth-stage companies where speed, scale, reliability, and operational discipline must be balanced.


For residents of San Francisco/Bay Area, CA or New York, NY the annual salary range for this role is $150,000-$220,000 + annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, annual bonus, and other factors.

About Forge Global

Forge Global is a financial services company that provides a marketplace for pre-IPO investments. The company was founded in 2014 and is headquartered in San Francisco, California. Forge Global's platform allows investors to buy and sell shares in private companies, providing liquidity to early investors and employees. The company has partnerships with over 200 private companies and has facilitated over $10 billion in transactions. Forge Global's mission is to democratize access to private markets and provide investors with opportunities for growth and diversification.
Learn more about Forge Global
Size
200 employees
Market Cap
$260.4 million
Industry
NASDAQ

Similar Jobs

More Jobs at Forge Global

More Information Technology Jobs

Find similar Manager, Site Reliability Engineer jobs: