Classification: Admin/Prof
Exemption Status/Test: Exempt/Computer
Job Grade: 5
Department: Information Technology
Reports to: Department Director
Job Goal: Design, implement, and continuously improve scalable, secure, and highly available infrastructure and deployment pipelines that enable rapid software delivery and reliable system performance. Drive automation across development and operations processes, leverage cloud technologies to optimize cost and efficiency, and collaborate closely with cross-functional teams to enhance system reliability, accelerate innovation, and support seamless, high-quality product releases.
Qualifications: Education: - Bachelor's degree within a technical field such as Computer Science or Information Technology, preferred
Experience: - 3 years of experience in DevOps, SRE, or related engineering roles
- Strong experience with AWS cloud platform (EC2, EKS, VPC, IAM, Auto Scaling Groups, Route53, Lambda, CloudWatch)
- Hands-on experience with ArgoCD or other GitOps tools for continuous deployment
- Experience with Helm charts for packaging and deploying Kubernetes applications
- Solid understanding of AWS IAM and IRSA (IAM Roles for Service Accounts) for secure pod-level permissions
- Experience with CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, CircleCI, etc.)
Special Knowledge and Skills: - Deep knowledge of Kubernetes including deployment strategies, resource management, and cluster operations
- Proficiency with containerization (Docker, container registries, multi-stage builds)
- Strong Infrastructure-as-Code skills with Terraform for provisioning and managing AWS resources
- Knowledge of MongoDB deployment and management in containerized environments
- Familiarity with Linux systems, networking, and security fundamentals
- Strong communication and collaboration skills
- Problem-solving mindset with a focus on reliability and scalability
- Ability to work in fast-paced, agile environments
- Ownership mentality and proactive approach to automation
Major Responsibilities - Build and maintain CI/CD pipelines to support automated testing, integration, and deployments.
- Manage cloud infrastructure using Infrastructure-as-Code tools such as Terraform, CloudFormation, or Pulumi.
- Monitor, troubleshoot, and optimize system performance across distributed applications and microservices.
- Implement automation for configuration management, provisioning, and workflow orchestration using tools like Ansible, Chef, Puppet, or similar.
- Enhance system reliability & availability through observability, log management, alerting, and incident response best practices.
- Collaborate closely with development teams to streamline build processes, improve deployment strategies, and ensure production readiness.
- Maintain containerized environments using Docker and orchestration tools (Kubernetes, ECS, EKS).
- Ensure security best practices in infrastructure, pipelines, secrets management, and access control.
- Document systems, standards, and procedures to improve maintainability and knowledge sharing.
Supervisory Responsibilities: None
Physical Demands/Environmental Factors/Mental Demands: Frequent use of standard office equipment; prolonged sitting; occasional bending/stooping, pushing/pulling, and twisting; repetitive hand motions (keyboarding and use of mouse); occasional light lifting and carrying (less than 15 pounds); may work prolonged and irregular hours; work with frequent interruptions; maintain emotional control under pressure.