Forge Global

Manager, Site Reliability Engineer

Forge Global$150K — $220K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years in Site Reliability Engineering, DevOps, or similar roles.
  • 10+ years of software engineering or cloud operations experience.
  • Bachelor's degree in Computer Science, Engineering, or related field.
  • Experience with large-scale cloud infrastructure and distributed systems.
  • Hands-on expertise in monitoring, incident response, and production support.
  • Familiarity with CI/CD, infrastructure automation, and operational tooling.
  • Strong communication and influencing skills across diverse teams.

Responsibilities

  • Lead the Site Reliability Engineering team at Forge for system availability.
  • Implement incident management processes, including mitigation and learning.
  • Enhance observability infrastructure alongside Platform Engineering.
  • Improve monitoring and alerting practices for quicker detection.
  • Promote reliability best practices like service ownership and disaster recovery.
  • Contribute to technical design and automation across the team.
  • Collaborate with teams to resolve production issues and boost reliability.

Benefits

  • Collaborative and creative engineering culture.
  • Opportunity to drive best practices in a growing team.
  • Focus on professional development and team health.
  • Engagement with cross-functional teams for holistic impact.
  • Chance to lead initiatives in a regulated infrastructure environment.
Full Job Description
The Role:

As an engineering organization, we pride ourselves on engineering as a creative activity. Engineering managers enable engineers to do their best work by maintaining a culture and environment where engineers can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge's SRE team responsible for keeping Forge systems highly available for customers, while partnering closely with Platform, Engineering, Security, Compliance, and Product teams to improve reliability, observability, incident response, and operational maturity. This is an opportunity for a hands-on technical leader who can coach engineers, improve production operations, and help Forge build and run secure, scalable, and highly reliable products.

Responsibilities:
  • Manage Forge's Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.
  • Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.
  • Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.
  • Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.
  • Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.
  • Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.
  • Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.
    • Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.
    • Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.

Qualifications:
  • 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.
  • 10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.
  • Bachelor's degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.
  • Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.
  • Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.
  • Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
  • Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.

Preferred Qualifications:
  • Experience in FinTech, financial services, or another regulated industry.
  • Experience with AWS and/or Azure cloud platforms.
  • Familiarity with Kubernetes, container platforms, infrastructure-as-code, Terraform, Ansible, or similar automation tooling.
  • Experience with observability platforms such as Datadog, CloudWatch, or similar tools.
  • Experience improving developer experience through paved-road platforms, standardization, and self-service infrastructure capabilities.
  • Experience supporting growth-stage companies where speed, scale, reliability, and operational discipline must be balanced.


For residents of San Francisco/Bay Area, CA or New York, NY the annual salary range for this role is $150,000-$220,000 + annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, annual bonus, and other factors.

About Forge Global

Forge Global is a financial services company that provides a marketplace for pre-IPO investments. The company was founded in 2014 and is headquartered in San Francisco, California. Forge Global's platform allows investors to buy and sell shares in private companies, providing liquidity to early investors and employees. The company has partnerships with over 200 private companies and has facilitated over $10 billion in transactions. Forge Global's mission is to democratize access to private markets and provide investors with opportunities for growth and diversification.
Learn more about Forge Global
Size
200 employees
Market Cap
$260.4 million
Industry
NASDAQ

Similar Jobs

More Jobs at Forge Global

More Information Technology Jobs

Find similar Manager, Site Reliability Engineer jobs: