Bachelor's degree in Computer Science, Software Engineering, or related field, or equivalent experience
5+ years in a DevOps engineering role with significant Azure experience
Proven experience designing and implementing large-scale Azure solutions
Hands-on experience with AKS in production environments
Advanced Terraform skills for module design and multi-environment deployments
Strong problem-solving abilities and solution architecture skills
Excellent communication and collaboration skills
Responsibilities
Build complex, highly available cloud infrastructure on Azure and Kubernetes (AKS)
Design and maintain production-grade AKS clusters with advanced networking and lifecycle management
Own infrastructure as code using Terraform, managing workspaces and state
Design and implement CI/CD pipelines ensuring quality and security
Establish monitoring and observability strategies across services
Identify and mitigate risks, leading troubleshooting for complex issues
Implement high-level security practices and manage secrets effectively
Maintain multi-region high availability and disaster recovery procedures
Benefits
Work in a strong community with top professionals in a friendly environment
Engage in large-scale projects with global impact
Access tailored learning opportunities including workshops and certifications
Explore diverse domains through internal mobility
Receive comprehensive healthcare and insurance benefits
Full Job Description
Job Description
About the role:
As a Senior DevOps Engineer, you'll become a part of a cross-functional development team engineering experience of tomorrow. We are looking for a talented Senior DevOps Engineer to support our engineering work, with a strong focus on the underlying infrastructure that we need to get from source code to production, plus keep it happy once it's there. We need someone who understands software development and enjoys working with developers on all the things necessary to improving, deploying, monitoring, and operating production services.
Responsibilities:
Build complex, highly available, and cost-optimized cloud infrastructure solutions on Azure and in Kubernetes (AKS)
Design, build, and maintain production-grade AKS clusters, including private cluster networking, node pool sizing and autoscaling, workload identity, ingress, certificate management, and cluster upgrade lifecycle
Own infrastructure as code end to end in Terraform: author and refactor reusable modules, manage per-environment workspaces and state, and drive changes through pull request, plan review, and apply
Design and implement advanced CI/CD pipelines ensuring quality, security, and efficiency, including container image build/publish and progressive promotion from lower environments to production
Establish comprehensive monitoring and observability strategies across cluster, application, and Azure platform services
Proactively identify and mitigate risks, leading troubleshooting efforts for complex issues across compute, networking, identity, and data layers
Implement high-level security practices, integrating vulnerability scanning and threat modeling, and manage secrets through a centralized secrets store with identity-based (rather than credential-based) access
Maintain and exercise multi-region high availability and disaster recovery, including failover and failback procedures for Kubernetes workloads and managed databases
Work with application engineering to cultivate operational standards
Requirements:
We know that sometimes, you can't tick every box. We would still love to hear from you if you think you're a good fit!
Bachelor's degree in Computer Science, Software Engineering, or a related field, or equivalent work experience
5+ years of experience in a DevOps engineering role with significant Azure experience
Proven track record of designing and implementing large-scale Azure solutions
Hands-on experience running AKS in production, not just in development or proof-of-concept environments - including private clusters, cluster upgrades, autoscaling, and incident response on live workloads
Advanced, production-level Terraform experience: module design, state and workspace management, and multi-environment/multi-subscription deployments
Demonstrated ability to solve complex technical problems and architect robust solutions
Cloud: Deep knowledge of Azure services, architectures, and design patterns, including Entra ID, RBAC, managed identities and workload identity federation, Key Vault, and subscription/landing zone structure
Infrastructure as Code (IaC): Advanced skills in Terraform, or other IaC tools
CI/CD: Strong experience with pipeline-as-code (Azure Pipelines, GitHub Actions, or equivalent), service connections and federated pipeline authentication, and GitOps-style continuous delivery to Kubernetes (Argo CD, Flux, or similar)
Programming and Scripting: Strong scripting skills (PowerShell, Bash, Python) and proficiency with one or more programming languages (e.g., Go, C#/.NET, Node.js)
Containerization and Microservices: Extensive experience with Docker, Kubernetes (AKS), Helm, container registries, and microservices architectures
Robust Networking: In-depth knowledge of networking concepts, Azure virtual network and subnet design, private endpoints and private DNS zones, hybrid connectivity (site-to-site VPN, network virtual appliances/firewalls), and network security
Data and Messaging Services: Familiarity with Azure managed data services such as SQL Managed Instance (including failover groups), Cosmos DB, Redis/managed cache, Storage, and Functions
Observability: Experience with Azure Monitor/Log Analytics and KQL, plus at least one third-party observability platform (Datadog, Grafana, or similar)
Resilience: Experience designing and testing multi-region redundancy, RTO/RPO-driven DR runbooks, and failover automation
Security Focus: Strong understanding of cloud security principles, penetration testing, and compliance standards (e.g., SOC 2, ISO 27001, PCI DSS)
Multi-cloud: ScriptCycle's footprint spans both Azure and Google Cloud, with an active workload migration track. Exposure to GCP and GKE (including Autopilot) and to cross-cloud networking and identity is a strong plus, though Azure is the primary focus of this role; experience supporting a GCP migration
What's in it for you?
Strong community: Work alongside top professionals in a friendly, open-door environment
Growth focus: Take on large-scale projects with a global impact and expand your expertise
Tailored learning: Boost your skills with internal events (meetups, conferences, workshops), Udemy access, language courses, and company-paid certifications
Endless opportunities: Explore diverse domains through internal mobility, finding the best fit to gain hands-on experience with cutting-edge technologies
Care: Healthcare, Basic Life Insurance, Short and Long-term disability insurance according to the Company's Benefit Plans
Interested already? We would love to get to know you! Submit your application. We can't wait to see you at Ciklum.