We are seeking a highly experienced Principal Site Reliability Engineering (SRE) leader to drive reliability, resilience, and secure software engineering practices across all critical applications within the Digital Consumer Engineering organization using applied AI.
In this role, you will operate at an enterprise and portfolio level, shaping reliability strategy, influencing architectural decisions, and enabling engineering teams to build and operate highly resilient, secure, and scalable platforms. You will play a critical role in advancing operational excellence, strengthening platform resilience, and supporting long-term business growth.
Primary Responsibilities:- Define and drive enterprise-wide SRE strategy, standards, and operating models across digital consumer platforms
- Establish and champion a culture of reliability engineering, security by design, and operational excellence
- Influence architecture, platform design, and investment decisions to improve reliability, scalability, and resilience at scale
- Partner with senior leaders across engineering, security, and operations to align reliability priorities with business outcomes
- Identify and incubate AI-driven capabilities to advance reliability, observability, automation, and proactive risk management
- Provide technical leadership and governance for reliability practices across systems including SLO frameworks and availability targets
- Drive enterprise-wide adoption of automation and self-healing systems
- Lead proactive risk management including threat modeling and resilience testing
- Establish reliability metrics and reporting frameworks to drive measurable improvements
- Partner with enterprise security leadership to embed security engineering practices into platform design
- Drive proactive threat detection and response strategies leveraging automation and AI
- Influence and standardize secure development and operational practices
- Lead cross-functional initiatives to strengthen preventative security posture
- Partner with platform teams to define scalable and resilient infrastructure strategies
- Drive adoption of infrastructure-as-code and automation practices
- Establish operational maturity and continuous improvement frameworks
Youll be rewarded and recognized for your performance in an environment that will challenge you and give you clear direction on what it takes to succeed in your role as well as provide development for other roles you may be interested in.
Required Qualifications:- Undergraduate degree in applicable area of expertise or equivalent experience
- 8+ years of relevant experience in SRE, software engineering, or infrastructure engineering
- 5+ years of experience operating at enterprise scale influencing strategy across multiple teams
- 4+ years of experience with in distributed systems, cloud infrastructure, and observability
- Proven track record of driving large-scale reliability or security transformation
- Demonstrated solid leadership and cross-functional influence skills
Preferred Qualifications:- Advanced degree
- Experience leading enterprise transformation initiatives
- Experience with AI/ML applied to operations or security
- Vendor/platform strategy experience
Pay is based on several factors including but not limited to local labor markets, education, work experience, certifications, etc. In addition to your salary, we offer benefits such as, a comprehensive benefits package, incentive and recognition programs, equity stock purchase and 401k contribution (all benefits are subject to eligibility requirements). No matter where or when you begin a career with us, youll find a far-reaching choice of benefits and incentives. The salary for this role will range from $134,600 - $230,800 annually based on full-time employment. We comply with all minimum wage laws as applicable.