The DCEO Cluster Manager position encompasses leadership responsibility for multiple data center facilities and their corresponding infrastructure. The successful candidate will direct a substantial organizational structure, providing oversight to Area Managers and Facility Managers while ensuring operational excellence across multiple data center locations.
This position requires demonstrated leadership capabilities and a track record of exceptional performance in managing complex technical operations. The role is integral to maintaining Amazon's world-class data center infrastructure and supporting our continued growth and innovation.
The ideal candidate for this role will have a strong understanding of data center/mission critical MEP infrastructure. You will be the single point of contact for all facility related issues and work as the technical resource to support Regional Cluster Manager. You will solely be responsible for maintaining 100% service uptime in your Cluster. You have a very high-level view of your organization, but you will need to dive into detail as needed. You will also be interacting with the global team daily by representing the Data Center Engineering Operations team.
Key job responsibilities
- Hiring, managing, and developing the operations management team including facility managers, area managers, chief engineers, and facility technicians.
- Establish performance benchmarks, conduct analyses, and prepare reports on all aspects of the critical facility operations and maintenance
- Responsible for the on-site management of 24x7 shift technicians, senior shift technicians, sub-contractors and vendors, ensuring that all work performed is in accordance with established practices and procedures.
- Build and maintain sustainable organizational structure.
- Work with business development, real estate, engineering and construction team to forecast staffing and maintain.
- Mentor and develop employees.
- Hiring strategy and participation in recruiting events.
- Operation and maintenance of mechanical, electrical, and controls systems for Amazon data centers include preventive maintenance, corrective maintenance, and change management.
- Manage all facility related repairs in a timely manner
- Daily/weekly/monthly/Data Center Engineering Operations meetings and reporting.
- Run weekly availability meetings and report outcomes.
- Liaise with a global team on global initiatives as well as process and procedures, implement those initiatives locally in Atlanta.
- Vendor management of colocation data center services providers to meet or exceed contracted performance SLAs.
- Safety, security, and availability of incident response, incident management, incident resolution, and root cause analysis.
- Cost management including OPEX and CAPEX associated with all data centers you oversee.
- Continuous improvement of operational processes, procedures, methods, and tools.
- Lead and manage energy efficiency initiatives in your Cluster.
- Support capacity planning and management and actively involved in ongoing construction activities.
A day in the life
Your day will be a blend of strategic oversight and hands-on problem-solving. You'll navigate complex facility challenges, coordinate with cross-functional teams, and ensure our data centers operate with precision and reliability. From monitoring critical systems to managing project timelines, every moment will be an opportunity to make a significant impact.
BASIC QUALIFICATIONS
- Bachelor's degree in Electrical Engineering, Mechanical Engineering, or a related field
- Experience hiring, developing, and managing high-performing technical teams
- 5+ years of managing managers experience
- 10+ years of mission critical facilities experience
- In depth knowledge of Data Center Facilities such as generators, chillers, cooling towers, air handling units, UPS, electrical sub distribution systems, fire detection and suppression systems, cable reticulation systems
PREFERRED QUALIFICATIONS
- Experience with process improvement techniques such as Kaizen, Lean Manufacturing or Six Sigma
- Experience with large-scale technical operations or large-scale compute farms
- Broad knowledge of information technology infrastructure domains such as computing server platforms, storage server platforms, server components, network devices, technologies and architectures, IT service delivery principles and best practices.
- Mission Critical facility management experience for a large enterprise or large Colocation provider
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.
USA, MS, Canton - 137,900.00 - 229,100.00 USD annually