As an Infrastructure Engineer at the Senior Software Engineer I level, you are a proficient infrastructure and platform engineer who is familiar with the patterns and practices of Infrastructure-as-Code and container orchestration. This person contributes to the team's consistent, reliable delivery of the infrastructure that supports the platform's iterative value to clients. The focus of this role is on provisioning, deploying, and operating infrastructure - including cloud resources, Kubernetes clusters, and CI/CD pipelines - rather than on writing application-layer code. This person collaborates with engineering teams to ensure infrastructure is secure, scalable, and reliably deployed across environments.
About the role- Provision and maintain cloud infrastructure using Terraform, including authoring and consuming versioned shared modules (e.g. IAM roles, Lambda layers, RDS cluster parameter groups, Secrets Manager secrets, security groups) and supplemental code written in TypeScript
- Manage and evolve Kubernetes (EKS) clusters, Helm charts, and Istio service-mesh configuration across environments
- Build and maintain CI/CD pipelines (GitHub Actions) supporting trunk-based merges to main
- Support provisioning and lifecycle management of cloud-hosted environments
- Manage secrets via AWS Secrets Manager and enforce least-privilege IAM policies across services and pipelines
- Participate in infrastructure code reviews, providing and incorporating feedback to improve reliability and maintainability
- Contribute to incident resolution and postmortem reviews for infrastructure and platform issues
- Produce clear infrastructure documentation and share knowledge with both technical and non-technical audiences
About you Minimum Qualifications
- Bachelor's degree in Computer Science, Information Systems, or a related field, or equivalent practical experience
- Experience provisioning and managing cloud infrastructure using Infrastructure-as-Code (Terraform preferred, Cloud Formation)
- Working knowledge of Kubernetes and Helm for packaging, configuring, and deploying containerized workloads
Preferred Qualifications
- Minimum of 2 years of experience in an infrastructure, platform engineering, DevOps, or Site Reliability Engineering role
- AWS certification (e.g., AWS Certified Solutions Architect or AWS Certified SysOps Administrator)
- Kubernetes certification (CKA or CKAD)
- Experience with Istio or another service mesh technology
- Experience with Aurora Postgres or other managed relational database services
- Experience with EKS or other container orchestration platforms
- Experience authoring and maintaining reusable, versioned Terraform modules
- Experience with observability and monitoring tooling (e.g. CloudWatch, Dynatrace, Splunk)
- Experience in capacity planning and dynamic infrastructure scaling
- Experience implementing different deployment strategies (rolling, blue/green, canary, etc.)
- Experience managing infrastructure with a major cloud provider (AWS preferred)
- Experience building and maintaining CI/CD pipelines (GitHub Actions or equivalent)
- Basic proficiency with Node.js and TypeScript for authoring infrastructure tooling scripts
- Strong troubleshooting skills, including the ability to reason about distributed systems and infrastructure dependencies
- Effective written and verbal communication skills, including the ability to document infrastructure decisions and processes
- Familiarity with Agile practices
What you'll get Our team members fuel our strategy, innovation and growth, so we ensure the health and well-being of not just you, but your family, too! We go above and beyond to give you the support you need on an individual level and offer all sorts of ways to help you live your best life. We are proud to offer eligible team members perks and health benefits that will help you have peace of mind. Simply put: We've got your back. Check out our full list of Benefits and Perks.
On-Call Expectations This role may include participation in an on-call rotation to support production systems and ensure service reliability. On-call responsibilities may include coverage during nights and weekends. If applicable, frequency and scheduling will be determined by team needs and communicated accordingly.