Job Title: Senior Compute Infrastructure EngineerOverview / Summary The role will lead compute-layer discovery and assessment activities, contribute to the consolidated risk register and remediation roadmap, and transition into a lead technical role within the ongoing managed services model. The engineer will support the client's transformation from an SOP-based operating model to a self-healing infrastructure model while working across hospital campuses, ambulatory practices, remote clinics, and vendor-hosted application dependencies.
Key Responsibilities - Lead compute-layer discovery activities across physical servers, hypervisors, virtual machines, and supporting infrastructure in coordination with Network Discovery and Cloud Connectivity workstreams.
- Operate discovery and inventory tools including Device42, Lansweeper, ServiceNow Discovery, and comparable platforms to produce a standardized configuration item inventory.
- Assess virtualization platforms including VMware vSphere, Microsoft Hyper-V, Nutanix AHV, or comparable technologies, covering cluster sizing, resource utilization, HA/DRS configuration, snapshot posture, and management-plane health.
- Perform infrastructure assessments in healthcare, financial services, or other regulated environments.
- Review healthcare clinical application hosting patterns, including EMR/EHR hosting environments, PACS/imaging archive dependencies, and HL7/FHIR integration engine hosting.
- Review server hardware lifecycle, firmware currency, BIOS/BMC posture, warranty status, and end-of-life/end-of-support risks.
- Assess operating system posture for Windows Server, Linux (RHEL, Ubuntu), and legacy platforms, including patch currency, support lifecycle, and standardized image alignment.
- Assess storage-related compute dependencies, including SAN/NAS connectivity, backup posture, and disaster recovery readiness at the hypervisor layer.
- Review clinical application hosting environments for server-level dependencies, performance, capacity, and platform standardization.
- Contribute compute-specific findings to the consolidated risk register based on clinical impact and operational feasibility.
- Serve as a senior technical resource within the NOC Managed Services model, providing Tier 3 escalation support for complex compute and virtualization issues.
- Execute compute change management activities, including VM lifecycle management, hypervisor upgrades, cluster rebalancing, patching, and server refresh activities.
- Support the transition to a self-healing infrastructure model through compute-layer automation using PowerShell, Ansible, Terraform, runbook creation, and monitoring improvements.
- Manage vulnerability reviews and patch planning for servers, hypervisors, and management-plane components.
- Collaborate with the client's Infrastructure Manager, Associate NOC Manager, and clinical application owners on change planning and clinical-window coordination.
- Contribute to weekly, monthly, and quarterly governance reporting through the Service Delivery Manager.
- Mentor junior compute engineers and NOC staff.
Required Qualifications - 13-14 years of experience.
- Deep technical knowledge and hands-on experience with Server Administration, including Active Directory, Group Policy, DNS, DHCP, WSUS/SCCM, and Windows Failover Clustering.
- Deep expertise in virtualization using VMware vSphere (ESXi, vCenter, HA/DRS), Microsoft Hyper-V, Nutanix AHV, or comparable platforms, including cluster design, capacity planning, and operational troubleshooting.
- Experience with discovery and inventory tools such as Device42, Lansweeper, ServiceNow Discovery, or comparable platforms.
- Experience performing infrastructure assessments in healthcare, financial services, or other regulated environments.
- Familiarity with healthcare clinical application hosting patterns, including EMR/EHR, PACS/imaging archive, and HL7/FHIR integration engine hosting.
- Experience with backup platforms such as Veeam, Commvault, Rubrik, Cohesity, or comparable technologies.
- Experience with compute-layer monitoring platforms such as SolarWinds SAM, Nagios, Datadog, ManageEngine OpManager, or comparable solutions.
- Experience with hyperconverged infrastructure, including Nutanix, VxRail, HPE SimpliVity, or comparable platforms.
- Strong experience with VMware vSphere, vCenter, ESXi, and underlying infrastructure.
- Proficiency in PowerShell scripting for automation and administrative tasks.
- Experience with Windows 10/11 tuning for virtual desktop environments.
- Knowledge of load balancing, security policies, MFA integrations, and session management.
- Experience with Microsoft Entra hybrid identity patterns, including domain-joined servers, Entra Connect, and hybrid Kerberos/certificate flows.
- Strong analytical and problem-solving capabilities.
- Excellent communication and stakeholder management skills.
- Ability to work independently and manage high-severity incidents under pressure.
- Collaborative team-oriented mindset.
Preferred Qualifications - VMware Certified Professional - Desktop and Mobility (VCP-DTM) or higher.
- Experience with hybrid VDI solutions (on-premises and cloud, such as Horizon Cloud on Azure/AWS).
- Exposure to Azure Virtual Desktop (AVD).
- ITIL Foundation certification or knowledge of change, incident, and problem management frameworks.
- Red Hat RHCE or comparable Linux certification.
- Nutanix NCP or comparable HCI certification.
#LI-ST1 #Hiring #LI-Onsite