Principal Test Infrastructure Architect
This role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office.
Job Description:Job Family Definition:Designs, architects, develops, troubleshoots, and improves the physical and virtual infrastructure used to validate HPE's SSR product. Determines infrastructure compatibility, influences hardware and network design, and ensures test environments are scalable, reliable, supportable, and aligned with current and future product needs. Works closely with Solutions and Development engineers to create, maintain, and trouble shoot testing solutions and automations.
Management Level Definition:Contributions regularly and sustainably impact technical components of HPE products, solutions, or services. Applies advanced subject matter knowledge to solve complex infrastructure, networking, automation, and operational issues and is regarded as a subject matter expert. Provides technical leadership and partnership to functional and cross-functional teams. Exercises significant independent judgment to determine the best methods for achieving objectives and may provide technical direction, mentoring, and leadership to others.
Responsibilities:- Owns the strategy, architecture, standards, reliability, capacity, operational readiness, and long-term roadmap for the SSR test environments in Westford and Andover, Massachusetts, and the remotely supported performance environment in Quincy, Washington.
- Designs and evolves physical lab layouts, network topologies, compute and virtualization platforms, storage, connectivity, management networks, and supporting services.
- Anticipates future infrastructure requirements based on SSR product plans, hardware platforms, network speeds, testing needs, and engineering priorities.
- Leads troubleshooting and root-cause analysis across hardware, Linux, networking, virtualization, power, connectivity, automation, and test infrastructure.
- Drives automation for equipment onboarding, provisioning, configuration, reservation, utilization, monitoring, recovery, inventory, and reporting.
- Establishes and maintains trusted systems of record, including NetBox or comparable tools, for equipment inventory, ownership, location, network connectivity, addressing, configuration, and lifecycle status.
- Plans lab capacity, equipment allocation, lifecycle requirements, and future infrastructure investments.
- Coordinates equipment purchasing, receipt, installation, rack placement, cabling, configuration, documentation, maintenance, and retirement.
- Performs or coordinates hands-on infrastructure work in the Westford and Andover labs when physical access is required.
- Supports the Quincy performance environment remotely in partnership with on-site data-center personnel and supporting HPE teams. Travel to Quincy is not expected.
- Partners with Engineering, Quality, Performance, Release Management, Tools, Infrastructure, Security, and corporate IT organizations.
- Works closely with team members in the United States and India to enable shared support, operational handoffs, remote software maintenance, automation, monitoring, and off-hours recovery activities without establishing a 24x7 support requirement.
- Reviews infrastructure designs and operational activities for compliance with lab standards and provides feedback to improve reliability, scalability, security, and product quality.
- Maintains architecture diagrams, network maps, operating procedures, recovery plans, troubleshooting guides, and onboarding documentation.
- Provides technical leadership, guidance, and mentoring to engineers working with lab infrastructure, automation, networking, and operational readiness.
- Establishes practical standards that allow engineers to perform routine lab activities without making the role an unnecessary approval point or operational bottleneck.
Education and Experience Required:- Bachelor's or Master's degree in Computer Science, Computer Engineering, Information Systems, Information Technology, or equivalent practical experience.
- Typically 6-10 years of experience in infrastructure architecture, site reliability engineering, engineering laboratories, data-center environments, network infrastructure, systems engineering, or a related technical field.
- Experience designing, supporting, or leading multi-rack, multi-network engineering, development, or product-validation environments.
Knowledge and Skills:- Extensive knowledge of Ethernet switching and routing, VLANs, network segmentation, addressing, DNS, DHCP, optics, cabling, and network troubleshooting.
- Strong experience with Linux systems, physical servers, storage, virtualization, and hardware lifecycle management.
- Experience with infrastructure automation using Python, Bash, Ansible, APIs, or comparable technologies.
- Experience with monitoring, observability, health checks, alerting, operational reporting, and automated remediation.
- Experience architecting and integrating physical, virtual, and cloud-based infrastructure across multiple platform types.
- Experience with NetBox, OpenStack, VMware, KVM, Docker, or similar infrastructure technologies.
- Familiarity with the HPE Juniper Networking Session Smart Routing product line and its hardware, software, management, and testing environments.
- Familiarity with Jenkins, GitHub, GitHub Copilot, Robot Framework, Jira, and Confluence.
- Experience with performance, regression, solution, WAN, hardware, or release-validation testing environments.
- Excellent analytical, troubleshooting, and problem-solving skills.
- Strong architecture, documentation, planning, prioritization, and cross-functional leadership skills.
- Ability to effectively communicate infrastructure architectures, design proposals, operational risks, and investment recommendations to engineering teams and senior management.
- Ability to collaborate effectively with globally distributed teams and lead technical initiatives through expertise, influence, and partnership.
- Ability to work regularly from HPE facilities in Westford and Andover, Massachusetts, and perform hands-on work with physical lab infrastructure when needed.
What We Can Offer You:Health & WellbeingWe strive to provide our team members and their loved ones with a comprehensive suite of benefits that supports their physical, financial and emotional wellbeing.
Personal & Professional DevelopmentWe also invest in your career because the better you are, the better we all are. We have specific programs catered to helping you reach any career goals you have - whether you want to become a knowledge expert in your field or apply your skills to another division.
Unconditional InclusionWe are unconditionally inclusive in the way we work and celebrate individual uniqueness. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good.
Let's Stay Connected:#unitedstates
#networking
Job:Engineering
Job Level:TCP_04
The expected salary/wage range for this position is provided below. Actual offer may vary from this range based upon geographic location, work experience, education/training, and/or skill level.
- United States of America: Annual Salary USD 136,500 - 260,500 in Massachusetts
The listed salary range reflects base salary. Variable incentives may also be offered.
Information about employee benefits offered in the US can be found at https://myhperewards.com/main/new-hire-enrollment.html