The OpportunityThe Manager, DevOps is a hands-on leader who sets direction, builds durable processes, and develops the team that keeps our cloud platforms reliable, secure, and cost-efficient. You will manage and grow the team while staying close to the technology, and you will be measured on the systems you put in place: how we plan, how we release, how we observe our environments, and how the team executes. The ideal candidate has built high-functioning DevOps practices before and knows how to bring structure and accountability to a team without slowing it down.
What You Will Do - Own the DevOps roadmap and delivery commitments across all Supplier.io environments, serving as the single point of accountability for DevOps direction and priorities.
- Manage, coach, and develop the DevOps team and own hiring for the team as it grows, including defining roles, interviewing, and onboarding.
- Establish and enforce core operational processes, including change management, release management, environment promotion standards, incident response and postmortems, on-call coverage, and documentation and runbook standards.
- Bring legacy, undocumented infrastructure under version control and documentation, and own disaster recovery and business continuity planning, including defined recovery time and recovery point objectives for our critical systems.
- Bring planning discipline to the function, including sprint planning, estimation, and clear delivery commitments.
- Define and drive the observability strategy, delivering centralized visibility across all environments with clear standards for monitoring, logging, alerting, and APM, and owning the related tooling decisions.
- Own the cloud cost optimization program, including committed-use discounts, right-sizing of compute and databases, and reduction of logging and storage costs.
- Drive security and compliance hygiene in partnership with the security function, including least-privilege access, access reviews, key rotation, certificate automation, and vulnerability remediation.
- Stay hands-on: build and review CI/CD pipelines, Infrastructure as Code, and Kubernetes workloads alongside the team, and lead by example on technical standards.
- Partner directly with Data Engineering to own the infrastructure and tooling that keeps their pipelines running: Snowflake account and warehouse administration, Airflow orchestration infrastructure, and secrets and connection management for data pipelines.
- Coordinate with the STAC on infrastructure decisions, such as cloud architecture, Kubernetes, and CI/CD tooling, that affect the broader platform, so infrastructure choices stay aligned with company-wide architecture standards.
- Participate in the on-call rotation alongside the team, especially early on, to stay close to the systems and set the standard for response quality.
- Build a culture of automation, accountability, and continuous improvement across the DevOps team and the broader engineering organization.
- Partner with software development, data engineering, AI, and platform teams to keep environments consistent and delivery fast, safe, and reliable.
- Evaluate emerging tools and practices, and make pragmatic decisions about what to build, buy, and adopt.
What You will Need to Succeed: - 10+ years in DevOps, cloud engineering, or infrastructure roles with progressive responsibility, including 3+ years leading or managing engineers with hiring, coaching, and performance responsibility.
- A track record of building processes that stick: change and release management, incident response, planning and estimation practices, and documentation standards in a growing organization.
- Experience bringing legacy or undocumented infrastructure under management, including introducing Infrastructure as Code and disaster recovery planning where neither existed before.
- Strong hands-on proficiency with CI/CD platforms (Azure DevOps, GitHub, Jenkins, or similar), Infrastructure as Code (Terraform, Ansible, Helm), and containerization and orchestration (Docker, Kubernetes).
- Deep experience with Google Cloud Platform. AWS and Azure experience is beneficial.
- Experience defining and implementing an observability strategy across multiple environments, including monitoring, logging, alerting, and APM tooling decisions.
- Experience managing cloud spend, including committed-use or reservation strategies, right-sizing, and cost accountability.
- Excellent communication skills, including the ability to set clear expectations with engineers, influence peers across engineering, and present plans and results to executives.
- Judgment to balance keep-the-lights-on operations, project delivery, and longer-term automation investments.
Preferred Qualifications:- Experience building or managing distributed teams, including offshore execution teams.
- Experience consolidating infrastructure and operations across acquired platforms.
- Experience with security automation, compliance frameworks, and policy-as-code.
- Familiarity with CI/CD for data pipelines and AI/MLOps practices.
We do no accept unsolicited resumes from recruitment/search firms.