This position requires that the candidate selected obtain and maintain an active Reliability Status security clearance with the Government of Canada.
Key job responsibilities
Key job responsibilities:
- Oversee all aspects of the data center's critical physical infrastructure. Maintain high availability and performance while ensuring that all work performed within the site is done to high quality without impact to internal/external customers and maintain service level agreements.
- Manage on-site teams of 24x7 shift technicians, sub-contractors and vendors ensuring that all work performed is in accordance with established practices, procedures and safety protocols.
- Drive safety and security culture and compliance within the team.
- Hire, coach, and promote team members. Responsible for team's career development and performance management.
- Promotes healthy team culture and peak morale.
- Effectively and efficiently manage the operations budget and expenditures.
- Engage in improvement projects, often requiring reaching out to a variety of support teams, and drive them from conception to completion.
- Act as an escalation point for all facilities-related issues. Oversee operation and management of routine and emergency services on a variety of critical systems such as: switchgear, generators, UPS systems, power distribution equipment, chillers, cooling towers, computer room air handlers, building monitoring systems, etc.
- Utilize root cause analysis and troubleshooting/problem solving skills.
- Respond to out-of-hours emergency calls.
- Role may support more than one location/site.
- May assist in the build out of new facilities and assist in projects to increase current facility efficiency.
- Utilize Microsoft Excel, Word, Outlook and other basic administrative tools to perform day to day tasks.
Principales responsabilités du poste:
- Superviser tous les aspects de l'infrastructure physique critique du centre de données. Maintenir une haute disponibilité et performance tout en s'assurant que tous les travaux effectués sur le site sont réalisés avec une haute qualité sans impact sur les clients internes/externes et maintenir les accords de niveau de service.
- Gérer les équipes sur site de techniciens en rotation 24/7, sous-traitants et fournisseurs en s'assurant que tous les travaux sont effectués conformément aux pratiques, procédures et protocoles de sécurité établis.
- Promouvoir la culture de sécurité et de sûreté ainsi que la conformité au sein de l'équipe.
- Embaucher, encadrer et promouvoir les membres de l'équipe. Responsable du développement de carrière et de la gestion de la performance de l'équipe.
- Promouvoir une culture d'équipe saine et un moral optimal.
- Gérer efficacement le budget opérationnel et les dépenses.
- S'engager dans des projets d'amélioration, nécessitant souvent de faire appel à diverses équipes de soutien, et les mener de la conception à l'achèvement.
- Agir comme point d'escalade pour tous les problèmes liés aux installations. Superviser l'exploitation et la gestion des services routiniers et d'urgence sur divers systèmes critiques tels que : appareillage de commutation, générateurs, systèmes UPS, équipement de distribution électrique, refroidisseurs, tours de refroidissement, unités de traitement d'air, systèmes de surveillance des bâtiments, etc.
- Utiliser l'analyse des causes profondes et les compétences de dépannage/résolution de problèmes.
- Répondre aux appels d'urgence hors horaires.
- Le rôle peut couvrir plus d'un emplacement/site.
- Peut aider à la construction de nouvelles installations et participer à des projets visant à améliorer l'efficacité des installations actuelles.
- Utiliser Microsoft Excel, Word, Outlook et d'autres outils administratifs de base pour effectuer les tâches quotidiennes.
BASIC QUALIFICATIONS
- 2+ years of people management and team development experience
- 5+ years of engineering work managing large-scale services experience
- Experience in people management and team development
- Experience in engineering work, managing large-scale services
- Experience maintaining SLAs through the implementation of proactive issue detection and reporting
- Experience operating a mission-critical team or product
- High school or equivalent
PREFERRED QUALIFICATIONS
- 3+ years of work in a management position with 5 or more direct reports experience
- 5+ years of work in data centers with an emphasis on building and equipment operation experience
- Bachelor's degree in Electrical Engineering, Mechanical Engineering, or a related field
- Knowledge of the electrical and mechanical systems involved in critical data center operations including systems such as feeders, transformers, generators, switchgear, UPS systems, ATS units, PDU units, chillers, pumps, air handling units, and CRAC units
- Experience in a management position with 5 or more direct reports
- Experience working in data centers with an emphasis on building and equipment operation
The base salary range for this position is listed below. As a total compensation company, Amazon's package may include other elements such as sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon offers comprehensive benefits including health insurance (medical, dental, vision, prescription, basic life & AD&D insurance), Registered Retirement Savings Plan (RRSP), Deferred Profit Sharing Plan (DPSP), paid time off, and other resources to improve health and well-being. We thank all applicants for their interest, however only those interviewed will be advised as to hiring status.
CAN, AB, Airdrie - 120,300.00 - 157,500.00 CAD annually