This is a Lead Infrastructure Production Management & Reliability Engineering position at the Vice President level, which is part of the job family responsible for maintaining the stability and reliability of the organization's infrastructure systems, ensuring optimal performance and availability to support business operations.
What you'll do in the role:
- Design, deploy, and support large-scale network infrastructure spanning Data Centers, Campus Networks, WAN Backbone, Extranet, Internet Edge, and Multi-Cloud environments.
- Engineer high-performance network platforms utilizing Arista 7800 Series, Cisco Nexus, and Cisco Next-Generation 8800 Routing Platforms.
- Design and support advanced routing architectures leveraging BGP, OSPF, MPLS, Segment Routing (SR-MPLS), EVPN/VXLAN, and MPLS SR Traffic Engineering (SR-TE).
- Lead strategic infrastructure transformation, network modernization, hardware refresh, and technology migration initiatives.
- Provide Tier-3 technical leadership for complex production incidents, critical escalations, and enterprise-wide network challenges.
- Architect and support secure hybrid-cloud connectivity solutions across Microsoft Azure, Google Cloud Platform (GCP), and on-premises environments.
- Design and optimize low-latency network solutions supporting business-critical applications, market connectivity, and high-performance trading environments.
- Support enterprise Extranet and Market Data connectivity, including exchange connectivity, third-party provider integrations, and ultra-low-latency network designs.
- Engineer and maintain Perimeter and Internet Edge infrastructure, including ISP connectivity, routing policy management, carrier diversity, and DDoS mitigation architectures.
- Partner with service providers and security teams to enhance network resiliency through DDoS protection, traffic engineering, peering strategies, and edge security design.
- Drive Infrastructure-as-Code and network automation initiatives using Python, Ansible, Terraform, APIs, GitHub, and CI/CD methodologies.
- Develop engineering standards, reference architectures, and deployment frameworks to improve scalability, resiliency, and operational efficiency.
- Lead security remediation, vulnerability management, and compliance-driven infrastructure enhancement initiatives.
What you'll bring to the role:
- 12+ years of enterprise network engineering experience supporting large-scale, mission-critical environments.
- Strong hands-on expertise with:
- Arista 7800 Series
- Cisco Nexus Platforms
- Cisco 8800 Routing Platforms
- Campus, Data Center, WAN, and Internet Edge networking
- Deep experience with:
- EVPN/VXLAN
- Leaf-Spine Architectures
- MPLS VPN
- Segment Routing (SR-MPLS)
- MPLS Traffic Engineering (SR-TE)
- BGP, OSPF, IS-IS, MP-BGP
- Strong knowledge of Extranet, Perimeter, Internet Edge, ISP, Market Data, and low-latency networking architectures.
- Experience designing and operating hybrid-cloud connectivity across Azure, Google Cloud Platform, ExpressRoute, Cloud Interconnect, and private WAN environments.
- Advanced automation expertise with Python, Ansible, Terraform, GitHub, APIs, and Infrastructure-as-Code frameworks.
- Proven success delivering large-scale upgrades, infrastructure transformations, security remediations, and network modernization programs.
- Strong packet analysis and troubleshooting skills utilizing Wireshark, telemetry platforms, and advanced network analytics tools.
- Deep understanding of network resiliency, high availability, disaster recovery, fault isolation, and scalable network architectures.
- Excellent communication and collaboration skills with the ability to work across engineering, security, cloud, application, and vendor organizations.
Preferred Qualifications
- Experience in financial services, market data, electronic trading, or other latency-sensitive environments.
- Knowledge of DDoS mitigation platforms, ISP peering, traffic engineering, and carrier network architectures.
- Experience with cloud-native networking, observability platforms, and emerging AI-driven network operations technologies.
- Industry certifications such as CCNP/CCIE Enterprise, Arista ACE, Google Professional Cloud Network Engineer, or equivalent.
To learn more about our offices across the globe, please copy and paste https://www.morganstanley.com/about-us/global-offices into your browser.
Expected base pay rates for the role will be between $155,000 and $215,000 per year at the commencement of employment. However, base pay if hired will be determined on an individualized basis and is only part of the total compensation package, which, depending on the position, may also include commission earnings, incentive compensation, discretionary bonuses, other short and long-term incentive packages, and other Morgan Stanley sponsored benefit programs.
To learn more about our offices across the globe, please copy and paste https://www.morganstanley.com/about-us/global-offices into your browser.