This is a role within CGI's Engineering group. A successful incumbent is expected to (i) Architect and build scalable, fault-tolerant, and secure cloud infrastructure using a rich AWS service portfolio (EKS, EC2, RDS, S3, IAM, VPC, Route 53, CloudWatch, and more)., and (ii) Lead the design and implementation of Infrastructure-as-Code (IaC) using Terraform, PowerShell, and other automation frameworks, enabling rapid and repeatable environment provisioning. Requires 10+ years of professional experience in Site Reliability Engineering, Software Development, Cloud Infrastructure Architecture, or related fields, with a strong track record of leading and delivering complex, large-scale projects. Prior leadership or mentorship experience, driving team growth, technical excellence, and collaborative culture. and a minimum of a Bachelor's degree in computer science, Engineering, Information Technology, or a related technical discipline. Relevant certifications (e.g., AWS Certified Solutions Architect, AWS Certified DevOps Engineer, Kubernetes Administrator) are highly desirable.
Job Responsibilities:
- Architect and build scalable, fault-tolerant, and secure cloud infrastructure using a rich AWS service portfolio (EKS, EC2, RDS, S3, IAM, VPC, Route 53, CloudWatch, and more).
- Lead the design and implementation of Infrastructure-as-Code (IaC) using Terraform, PowerShell, and other automation frameworks, enabling rapid and repeatable environment provisioning.
- Drive GitOps excellence with Argo CD, enabling continuous delivery and deployment for Kubernetes workloads with confidence and speed.
- Build an industry-leading observability platform with Datadog, empowering teams with real-time insights through metrics, logs, traces, and synthetic monitoring.
- Own the architecture and enforcement of network security and network management across AWS and hybrid environments - ensuring our infrastructure is both resilient and secure.
- Define, measure, and optimize SLIs, SLOs, and SLAs that align reliability efforts directly with business impact.
- Build and refine automated incident response, runbooks, and proactive monitoring to minimize downtime and accelerate recovery.
- Collaborate closely with software engineering, product, and security teams to integrate reliability and security into the development lifecycle from day one.
- Influence capacity planning, cost optimization, and cloud strategy to support aggressive growth and innovation.
- Indirect Supervision to lead, mentor, and inspire a high-performing SRE and DevOps team committed to operational excellence
- Comply with health and safety guidelines and rules; managers should also ensure compliance across their teams.
- Protect Chamberlain Group's reputation by keeping information confidential.
- Maintain professional and technical knowledge by attending educational workshops, reading professional publications, establishing personal networks, and participating in professional societies.
- Contribute to the team effort by accomplishing related results and participating on projects as needed.
Job Requirements:
- Bachelor's degree in computer science, Engineering, Information Technology, or a related technical discipline.
- Relevant certifications (e.g., AWS Certified Solutions Architect, AWS Certified DevOps Engineer, Kubernetes Administrator) are highly desirable.
- 10+ years of professional experience in Site Reliability Engineering, Software Development, Cloud Infrastructure Architecture, or related fields, with a strong track record of leading and delivering complex, large-scale projects.
- Prior leadership or mentorship experience, driving team growth, technical excellence, and collaborative culture.
- Up to around 5% travel time
Knowledge, Skills, and Abilities:
- Proven expertise in architecting and managing cloud-native infrastructure, particularly on AWS.
- Demonstrated hands-on experience with Infrastructure-as-Code tools such as Terraform and scripting languages including PowerShell, Python, or Bash.
- Experience leading and implementing GitOps workflows with tools like Argo CD.
Preferred Job Requirements:
- Master's degree in computer science, Engineering, Information Technology, or a related technical discipline.
- One or more of AWS Certified Solutions Architect, AWS Certified DevOps Engineer, Kubernetes Administrator certifications
- Proficient in designing and maintaining observability solutions using Datadog or comparable platforms.
- Advanced knowledge of network security, VPC design, firewall management, and secure connectivity.
- Exceptional troubleshooting and incident management abilities, coupled with strong communication skills.
- Passion for mentoring teams and fostering a culture of innovation, reliability, and operational excellence.
- Experience with global and distributed cloud-based IoT platforms and devices.
Knowledge, Skills, and Abilities:
- Deep expertise in Terraform and other IaC frameworks with proven production experience.
- Mastery of Kubernetes architecture, deployment, and GitOps automation using Argo CD or equivalent.
- Strong programming and automation skills in PowerShell, Python, Go, Bash, or similar languages.
The pay range for this position is $129,700.00 - $226,900.00; base pay offered may vary depending on a number of factors including, but not limited to, the position offered, location, education, training, and/or experience. In addition to base pay, also offered is a comprehensive benefits package and 401k contribution (all benefits are subject to eligibility requirements). This position is eligible for participation in a short-term incentive plan subject to the terms of the applicable plans and policies.