We are seeking a Platform engineer to design, deploy, administer, secure, and support global virtualization platforms using VMware and OpenShift Virtualization (OSV). This role owns the full technology lifecycle, drives automation and observability, manages capacity and performance, and provides advanced troubleshooting and platform consulting for enterprise applications.
ResponsibilitiesVirtualization Hosting & Platform Engineering- Engineer, deploy, administer, secure, and manage the full lifecycle of global VMware and OpenShift Virtualization (OSV) platforms.
- Design and support large-scale virtualization environments, including HPE Synergy/ProLiant iLO and firmware testing.
- Develop and maintain operating procedures, monitoring and logging standards, disaster recovery processes, and security guidelines.
Capacity Management- Plan and forecast compute, VM, memory, storage, and network capacity.
- Analyze utilization trends and recommend scaling, consolidation, and optimization strategies.
Automation & Platform Efficiency- Automate VMware/OSV configuration, VM management, auditing, remediation, and ticketing workflows using scripts and playbooks.
- Apply SRE practices to improve platform reliability, performance, and operational efficiency.
- Manage RBAC, namespaces, and CPU, disk, and storage quotas.
Observability & Troubleshooting- Implement end-to-end monitoring, logging, tracing, and alerting using Dynatrace, RHACM, Prometheus, and Grafana.
- Monitor VM health, performance, resource utilization, and potential security anomalies.
- Lead root cause analysis and resolve complex issues across the global compute environment.
Solution Design & Consulting- Consult with application teams and design custom OSV clusters and VM solutions for critical or specialized workloads.
Documentation & Support- Maintain technical documentation and customer-facing self-service content.
- Provide L1-L3 support for Operations teams.
- Participate in monthly after-hours and weekend support.
Job Requirements
Details:
Skills Required- Scripting, Automation, Kubernetes, Root cause analysis, Cloud architecture, IT solutions, GitHub, Cloud infrastructure, Change management, Technical analysis, Development, Tekton, Utilization management, VMware, VMware ESX Servers, Platform support, Infrastructure architectures
Skills Preferred<
- Ansible, Google Cloud Platform (GCP), Dynatrace, PowerShell, Access controls, Python, Information security, Automation, Artificial intelligence and expert systems
Experience Required- 10+ years of experience in IT, 8+ years of experience in development.
- Understanding of VMware and Kubernetes concepts.
- Experience with Linux administration and networking fundamentals.
- Proficiency in scripting languages for automation.
- Experience with monitoring tools and logging solutions.
- Understanding of virtualization concepts and technologies, such as KVM and VMware.
- Excellent problem-solving skills and the ability to troubleshoot complex issues across multiple layers of the technology stack.
- Knowledge of CI/CD pipelines and DevOps methodologies.
Education Required- Associate Degree, College Senior
Education Preferred- Certification Program, Bachelor's Degree
Pay Range:
$ 61.00 - $ 66.00