OverviewThe Multi-Cloud Engineer – SME serves as a senior technical authority and hands-on engineer for designing, building, integrating, automating, and sustaining capabilities across multiple cloud providers and on-premises environments. The role establishes technical patterns, resolves complex architecture and integration challenges, develops production-ready solutions, and advises government and program leadership on multi-cloud strategy.
Responsibilities
- Serve as the senior technical subject-matter expert and hands-on engineer for multi-cloud capabilities across Microsoft Azure, Amazon Web Services (AWS), and hybrid/on-premises environments.
- Design, build, configure, deploy, test, troubleshoot, and sustain cloud infrastructure, platform services, and integrations in production and non-production environments.
- Develop reference architectures, reusable infrastructure modules, engineering standards, and implementation guidance that enable portability, interoperability, resilience, and consistent security across cloud providers.
- Implement and maintain Infrastructure as Code (IaC) using tools such as Terraform, Bicep/ARM templates, AWS CloudFormation, Ansible, or equivalent automation frameworks.
- Build and maintain continuous integration/continuous delivery (CI/CD) pipelines for infrastructure, platform components, and cloud-native applications using approved DevSecOps tooling.
- Design and implement cross-cloud and hybrid connectivity, including virtual networks/VPCs, subnets, routing, firewalls/security groups, private connectivity, domain name system (DNS), load balancing, network segmentation, and connectivity to on-premises environments.
- Configure and troubleshoot identity, access management (IAM), privileged access, role-based access control (RBAC), identity federation, service identities, secrets management, and integration with enterprise identity providers.
- Deploy, configure, administer, upgrade, secure, and troubleshoot container platforms and Kubernetes environments, including cluster networking, ingress, storage, policy enforcement, workload deployments, and observability.
- Develop and maintain cloud-native application hosting patterns for virtual machines, containers, serverless services, managed databases, object storage, messaging, and API integration services.
- Implement automated configuration management, patching, vulnerability remediation, image hardening, baseline enforcement, and compliance reporting for cloud and hybrid infrastructure.
- Build and tune centralized logging, monitoring, alerting, tracing, dashboards, and service-level indicators across cloud and on-premises environments.
- Conduct root-cause analysis for complex infrastructure, networking, application-platform, identity, performance, and availability incidents; implement corrective and preventive actions.
- Lead hands-on engineering efforts for complex migrations, modernization initiatives, workload refactoring, data migration, and hybrid-cloud transformations.
- Evaluate cloud services and prototype solutions to recommend fit-for-purpose technologies based on mission requirements, performance, cost, security, and operational constraints.
- Analyze cloud performance, availability, capacity, resiliency, and cost; implement optimization measures, autoscaling, backup, disaster-recovery, and recovery-validation solutions.
- Produce and maintain build documentation, runbooks, scripts, configuration baselines, implementation plans, test procedures, architecture diagrams, and operational procedures.
- Perform peer reviews of infrastructure code, cloud configurations, technical designs, implementation artifacts, and operational documentation.
- Mentor cloud engineers through code reviews, technical demonstrations, troubleshooting support, and hands-on training.
- Brief senior government and program leadership on technical risks, tradeoffs, modernization options, implementation status, and recommended courses of action.
Qualifications
Required Qualifications
- Bachelor’s degree in a technical discipline or equivalent relevant experience.
- 15+ years of progressively responsible infrastructure and cloud engineering experience.
- Expert-level, hands-on experience designing, deploying, configuring, and troubleshooting Azure and AWS services in enterprise environments.
- Demonstrated ability to create, modify, test, and maintain IaC, automation scripts, and deployment pipelines using Terraform, Bicep/ARM, CloudFormation, Ansible, PowerShell, Python, Bash, or equivalent tools.
- Deep hands-on experience with cloud networking, including Azure Virtual Network, AWS Virtual Private Cloud, routing, VPN/private connectivity, DNS, firewalls, load balancers, network security controls, and hybrid connectivity.
- Deep hands-on experience implementing IAM, federation, RBAC, workload identities, privileged access, secrets management, and cloud security controls in Azure and AWS.
- Demonstrated experience deploying, administering, securing, and troubleshooting Kubernetes and container ecosystems, such as Azure Kubernetes Service, Amazon Elastic Kubernetes Service, OpenShift, or equivalent platforms.
- Experience operating CI/CD and DevSecOps pipelines and integrating security scanning, policy checks, automated testing, artifact repositories, and deployment controls.
- Practical experience configuring cloud monitoring, logging, alerting, incident response workflows, backup/recovery, high availability, and disaster recovery capabilities.
- Demonstrated experience leading complex technical initiatives while personally contributing to implementation, troubleshooting, migration, and sustainment activities.
- Strong understanding and applied experience with Zero Trust principles, cloud security architecture, resilience engineering, disaster recovery, and regulated-environment requirements.
- Ability to produce clear technical documentation, automate repeatable operational tasks, mentor engineers, and communicate complex technical issues to technical and nontechnical stakeholders.
- Ability to operate independently as a senior technical authority while collaborating across cybersecurity, networking, software, data, infrastructure, and mission teams.
- Must hold a CompTIA Security+ CE or other approved 8140 IAT Level II certification.
- Must hold at least one relevant cloud certification in either Azure or AWS.
- Must reside within a75-mile radius of Redstone Arsenal in Huntsville ALor be willing to relocate prior to starting date.
- Active Secret security clearance must be able to maintain the security clearance required for this position.
Preferred Qualifications
- Experience in DoW/DoD mission environments.
- Experience with Azure Government, AWS GovCloud, IL5/IL6, or comparable regulated environments.
- Cloud certifications such as AWS Solutions Architect/Engineer or Microsoft Azure Administrator/Engineer.
- Advanced Azure and AWS certifications; Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS), HashiCorp Terraform certification; or equivalent demonstrated expertise.
- Experience with Git-based engineering workflows, code review, artifact repositories, container registries, and automated release processes.
- Experience with FinOps practices, cloud cost allocation/tagging, budget monitoring, rightsizing, reserved-capacity strategies, and cloud economics analysis.
- Experience with platform engineering, internal developer platforms, service catalogs, policy-as-code, GitOps, and large-scale data-center modernization.
- Experience scripting and automating administrative tasks with Python, PowerShell, Bash, Go, or similar languages.
- Preferred experience with SPECTRO Cloud, including deployment, configuration, administration, and integration of SPECTRO Cloud capabilities within hybrid and multi-cloud environments.
- Preferred experience architecting, deploying, and integrating SPECTRO Cloud as part of secure hybrid and multi-cloud environments, including platform orchestration, workload management, automation, and integration with government mission systems.
Physical Demands
- The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job.
- While performing the duties of this job, the employee is regularly required to talk or hear. The employee frequently is required to stand; walk; handle or feel; and reach with hands and arms.
- The employee is occasionally required to sit; climb or balance; and stoop, kneel, crouch or crawl. The employee must be able to lift and/or move up to 10 pounds and occasionally lift and/or move up to 25 pounds.
- Specific vision abilities required by this job include close vision, distance vision, peripheral vision, depth perception and ability to adjust focus.
- Regular i3 hours are 8:00 a.m. - 5:00 p.m. Monday-Friday, however, additional hours may be required on occasion. Regular and punctual attendance is required.