Full Job Description
In the role of VCF Platform Engineer Manager, we'll count on you to:
- Directly manage VCF Platform engineers and/or leads; set clear expectations for delivery, accountability, collaboration, and professional growth.
- Own the platform roadmap and service catalog; prioritize the engineering backlog based on business value, risk, lifecycle requirements, and operational needs.
- Track and report platform KPIs, including availability, provisioning time, cost per VM, incident mean time to restore (MTTR), utilization, and service delivery performance.
- Serve as the senior escalation point for platform incidents, service degradation, capacity constraints, and cross-functional issue resolution.
- Coordinate with FinOps, cybersecurity, application, network, storage, and infrastructure teams through defined operational and governance interfaces.
- Lead the weekly planning cadence for the platform team, including backlog review, priorities, delivery status, dependencies, and risk management.
- Report platform health, cost, service performance, and strategic recommendations to infrastructure leadership.
- Ensure the private cloud platform is operated as a reliable, scalable, cost-effective service aligned to enterprise standards and business requirements.
- Drive continuous improvement in platform operations, engineering practices, lifecycle management, resiliency, and service consumption.
- Partner with architecture and engineering stakeholders to evaluate platform enhancements, modernization opportunities, and future-state capabilities.
- Support operational governance activities, including change review, incident review, problem management, capacity planning, and service performance reviews.
Preferred Qualifications
- VMware certification such as VCP-VCF, VCP-DCV, or other relevant VMware platform certifications.
- Experience operating infrastructure using a platform-as-a-product or service ownership model.
- Experience with FinOps, cost modeling, infrastructure unit economics, showback/chargeback, or cost optimization practices.
- Familiarity with hybrid cloud operating models and integration points between private cloud and public cloud services.
- Experience with automation and infrastructure management tools such as PowerCLI, Terraform, Ansible, vRealize/Aria, or similar orchestration platforms.
- Experience with service catalog management, self-service provisioning, and standard service offering design.
- Knowledge of resiliency, disaster recovery, backup, and high availability design principles for enterprise platforms.
- Exposure to cybersecurity and compliance requirements relevant to infrastructure platforms, including vulnerability remediation and secure configuration practices.
- Experience supporting multi-site or geographically distributed data center platforms.
- Prior experience building dashboards, executive reporting, and KPI visualization for operational leadership.
- ITIL certification or equivalent practical service management experience.
- Schedule & Presence: This on-site role supports 24/7 operations through real-time collaboration, standard shifts occur within a 6:00 AM - 6:00 PM window, Monday through Friday. Additionally, this position requires scheduled on-call flexibility and the ability to remain reasonably reachable during off-hours for critical business continuity.
Qualifications
Required Qualifications
- Bachelor's degree in Information Technology, Computer Science, Engineering, or a related field; or an equivalent combination of education and relevant experience.
- A minimum 7 years of progressive experience in infrastructure, private cloud, virtualization, or enterprise platform operations.
- A minimum 5 years of leadership experience managing technical teams, engineering backlogs, service delivery functions, or platform operations.
- Demonstrated experience administering, operating, or leading enterprise VMware vSphere / VMware Cloud Foundation (VCF) environments.
- Experience owning or managing shared infrastructure or platform services with accountability for availability, capacity, service quality, and cost performance.
- Strong working knowledge of IT service management (ITSM) processes, including incident, change, problem, request, and service level management.
- Proven experience in capacity planning, utilization analysis, performance monitoring, and infrastructure lifecycle planning.
- Experience defining, tracking, and reporting operational and service KPIs to technical and leadership audiences.
- Strong stakeholder management and communication skills, with the ability to align technical priorities to business objectives.
- Experience coordinating across multiple infrastructure disciplines such as compute, storage, networking, security, and operations.
- Demonstrated ability to make sound operational decisions during incidents, escalations, and competing priority scenarios.
- Experience leading planning cadences, work prioritization, and execution tracking for technical teams.