Job DescriptionSummary:The DevOps Engineer II supports engineering teams that build, deploy, operate, and scale ADT's mission-critical monitoring platforms across cloud and on-premise environments. This role helps implement reusable platform capabilities, CI/CD pipelines, infrastructure automation, observability solutions, and operational tooling that improve delivery speed, reliability, security, and developer experience. The ideal candidate is a hands-on engineer who can work with platform engineering, product engineering, architecture, security, QA, and SRE teams to support consistent deployments and reliable operations.
Duties and Responsibilities:Platform Enablement- Build and maintain shared platform capabilities used by engineering teams.
- Support infrastructure patterns for cloud-native and on-premise workloads.
- Create and improve deployment and operational tooling that supports developer productivity.
- Partner with engineering teams to reduce delivery bottlenecks and improve release execution.
- Help standardize deployment, monitoring, documentation, and operational practices across teams.
Cloud & Infrastructure Engineering- Implement infrastructure using Infrastructure as Code practices under guidance from senior engineers and architects.
- Support workloads running across Google Cloud Platform, hybrid cloud environments, and on-premise data centers.
- Assist with secure, scalable, and highly available infrastructure services.
- Support platform resiliency, disaster recovery readiness, and operational reliability improvements.
CI/CD & Software Delivery- Build, maintain, and improve CI/CD pipelines for application and platform teams.
- Automate build, test, security, and deployment workflows.
- Partner with developers and QA engineers to incorporate automated quality gates.
- Support release automation, deployment safety, rollback readiness, and repeatable delivery workflows.
Containerization & Platform Operations- Support containerized environments using Docker and Kubernetes.
- Assist with workload portability across cloud and on-premise platforms.
- Develop reusable deployment templates, operational scripts, and platform documentation.
- Maintain runbooks and operational standards for supported platform capabilities.
Observability, Reliability & Support- Implement and maintain monitoring, logging, tracing, and alerting solutions.
- Partner with SRE and engineering teams to improve system availability and operational health.
- Support incident response, root cause analysis, and post-incident improvement activities.
- Help develop automated remediation and operational response capabilities.
Security & Compliance- Partner with Security teams to implement secure infrastructure patterns and engineering workflows.
- Assist with infrastructure compliance, governance controls, and security validation automation.
- Integrate security scanning and validation into software delivery pipelines.
- Support compliance requirements for safety, monitoring, and regulated environments.
- Additional duties as assigned.
Skills and Competencies:- Working knowledge of DevOps, platform engineering, cloud infrastructure, CI/CD, and automation practices.
- Hands-on experience with cloud platforms, infrastructure automation, scripting, source control, and deployment tooling.
- Ability to work collaboratively with engineering, QA, security, SRE, architecture, and operations teams.
- Strong troubleshooting skills and ability to investigate issues across infrastructure, pipelines, deployments, and application runtime environments.
- Comfort working in hybrid cloud and on-premise environments.
- Clear written and verbal communication skills, including the ability to document runbooks, deployment steps, and operational procedures.
- Ownership mindset, willingness to learn, and focus on improving reliability, security, and developer experience.
Minimum Qualifications:- Bachelor's degree in Computer Science, Software Engineering, Information Systems, or related field, or equivalent practical experience.
- 5+ years of experience in DevOps, Platform Engineering, Site Reliability Engineering, Cloud Engineering, Systems Engineering, or related technical disciplines.
- Experience with CI/CD pipelines, build automation, deployment automation, or release workflows.
- Experience with cloud platforms, preferably Google Cloud Platform, or comparable cloud environments.
- Experience with scripting or automation using Python, Bash, Go, or similar technologies.
- Familiarity with containers, Kubernetes, Infrastructure as Code, networking, security, compute, storage, and identity concepts.
Preferred Qualifications:- Experience supporting production, hybrid cloud, or on-premise environments.
- Experience with observability platforms, logging, tracing, and alerting tools.
- Experience with automated testing, security scanning, quality gates, and deployment validation.
- Familiarity with mission-critical, high-availability, monitoring, safety, or regulated environments.
- Interest in developer experience, internal platform tooling, reliability engineering, and operational excellence.
Required Licensing or Certifications:- Relevant certifications in Cloud Technologies, DevOps, Kubernetes, Security, Agile Practices, or Site Reliability Engineering are preferred.
Working Conditions:- Hybrid work environment with a combination of office and remote work.
- Frequent collaboration with platform engineering, product engineering, QA, security, SRE, architecture, and operations teams.
- Work may involve production, pre-production, cloud, and on-premise monitoring platform environments.
- May require participation in release support, incident review, operational readiness, or deployment validation activities.
Travel:- Limited travel may be required for team collaboration, planning sessions, operational readiness activities, or business needs.