Federal Home Loan Bank of Des Moines

Manager Site Reliability Engineering

Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 2+ years in Site Reliability Engineering or a closely related field
  • 5+ years in infrastructure, operations, or DevOps
  • 2+ years of people management experience, with focus on developing technical talent
  • Strong foundation in systems administration, networking, and infrastructure
  • Proficiency in programming or scripting languages, preferably .NET or Python
  • Familiarity with monitoring and observability tools
  • Demonstrated ability to influence without direct authority

Responsibilities

  • Establish and implement SRE practices across the organization
  • Build and lead a team of 3-4 engineers transitioning to SRE
  • Create a culture promoting psychological safety and continuous improvement
  • Define and enforce Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
  • Lead optimization of monitoring and alerting tools
  • Participate in incident management and disaster recovery exercises
  • Advocate for reliability initiatives within the organization

Benefits

  • Highly competitive compensation and bonus packages
  • Comprehensive benefits program
  • 401(k) and pension plan
  • Five weeks of vacation for new employees and 11 paid holidays per year
  • Annual stipend for wellbeing activities through the Lifestyle Spending Account
  • Central downtown location with easy public transport access and scenic rooftop views
Full Job Description


What you'll do

We are building a new Site Reliability Engineering function and seeking a leader who can establish SRE practices across the organization while developing a team of engineers new to the discipline. This is a unique opportunity to shape how reliability engineering is practiced at FHLBank Chicago from the ground up.

The SRE team operates as a guiding and consultative partner to application and development teams rather than owning systems directly. Success in this role requires technical credibility, strong influencing skills, and the ability to drive change through collaboration and education rather than direct authority. This role is accountable for building and leading the SRE team, including hiring, performance management, coaching, and development of engineers transitioning into SRE practices. The manager establishes the SRE operating model (how SRE engages with application and development teams), ensures sustainable on-call and learning culture, and drives adoption of reliability standards through collaboration and influence.

How you'll make an impact
  • Shape how Reliability Engineering is practiced and enforced across the Bank through collaboration
  • Build deep relationships between IT and the greater organization in support of common goals
  • Provide direction and development guidance to a team of 3-4 members in this new space
  • Deepen further automation into application availability and reporting processes


What you can expect

Team Leadership and Development
  • Build and develop a team of engineers transitioning from traditional operations and systems administration backgrounds into SRE practices
  • Create psychological safety that enables learning, experimentation, and honest discussion of failures
  • Establish career development paths and growth opportunities within the SRE discipline
  • Foster a culture of blameless postmortems and continuous improvement


Reliability Engineering Practice
  • Define and implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budget policies across critical services
  • Establish pager budgets and on-call practices that are sustainable and effective
  • Lead tuning and optimization of monitoring, alerting, and observability tooling
  • Drive reduction of system disruptions through automation, tooling, and process improvement
  • Develop and maintain incident management processes, including severity classification and escalation procedures
  • Participate in FHLBank's Disaster Recovery process and testing, coordinating and executing regularly scheduled DR exercises


Consultation and Partnership
  • Participate in troubleshooting meetings and production incidents, providing expert guidance and recommendations
  • Partner with application owners, product owners, and development teams to improve system reliability
  • Ensure deep technical analysis is performed for significant reliability issues, and provide escalation support as needed; coach the team in translating findings into actionable recommendations.
  • Advocate for reliability investments and help teams prioritize reliability work against feature development
  • Build relationships that enable SRE to influence architectural and operational decisions without direct ownership


Organizational Leadership
  • Secure and maintain executive sponsorship and governance mechanisms required for SLO and error budget practices (including defined decision rights when reliability thresholds are breached).
  • Communicate the value and principles of SRE to leadership, helping secure sustained support and appropriate resource allocation
  • Develop metrics and reporting that demonstrate SRE impact on business outcomes
  • Navigate organizational dynamics to build credibility and trust for a new function
  • Align SRE practices with existing compliance, risk management, and regulatory requirements


What you'll bring
  • 2+ years of Site Reliability Engineering or directly-related primary function
  • 5+ years of experience in infrastructure, operations, DevOps
  • 2+ years of people management experience, with demonstrated ability to develop and grow technical talent
  • Strong technical foundation in systems administration, networking, and infrastructure
  • Proficiency in at least one programming or scripting language (.NET preferred, Python also valuable)
  • Experience with monitoring, observability, and alerting tools and practices
  • Proficiency with agentic AI tools such as Github Copilot, Claude Code, or Codex
  • Demonstrated ability to influence outcomes without direct authority
  • Strong written and verbal communication skills, including ability to explain technical concepts to non-technical stakeholders
  • Experience conducting or leading incident response and postmortem processes
  • Outstanding communication (verbal, written, and listening) skills
  • Proven ability to consistently navigate crucial conversations
  • Critical thinking - using logic and reasoning to identify the strengths and weaknesses of alternative solutions, conclusions or approaches to problems
  • Systems thinking - approaching situations and scenarios understanding they are a complex web of interdependencies between items and other systems that are often initially unclear
  • Ability to present ideas in business-friendly and user-friendly language
  • Attention to detail
  • Comfort with high levels of ambiguity and shared responsibility
  • Pleasant demeanor with others with a good-natured, cooperative attitude
  • Experience with Agile methods and concepts
  • Knowledge of cloud computing principles, specifically related to Amazon Web Services


Preferred Qualifications
  • Experience implementing SRE practices in an organization new to the discipline
  • Background in financial services or other regulated industries
  • Experience defining and implementing SLIs, SLOs, and error budget policies
  • Experience building or transforming teams through organizational change
  • Knowledge of ITIL, DevOps, or related frameworks


The Perks

At FHLBank Chicago, we believe in rewarding our high performing workforce. We offer a highly competitive compensation and bonus package, and access to a comprehensive benefits program designed to meet the needs of our employees. Our retirement program includes a 401(k) and pension plan. Our wellbeing program supports employees at work and in their personal lives: Our PTO plan provides five weeks of vacation for new employees and 11 paid holidays per year; our Lifestyle Spending Account provides an annual stipend for employees to support wellbeing activities; and our central downtown location at the Old Post Office provides easy access to public transportation and breathtaking views from our award-winning rooftop. Visit FHLBCbenefits.com for additional details about our benefits. Step into a brighter future with us.

Salary Range:

$125,825.00 - $221,275.00

The above represents the expected salary range for this job requisition. Ultimately, in determining your pay, we may also consider your experience, and other job-related factors. In addition to the base salary, we offer a comprehensive benefits package which can be found here: https://hrportal.ehr.com/fhlbc

About Federal Home Loan Bank of Des Moines

The Federal Home Loan Bank of Des Moines is a member-owned cooperative that provides funding solutions and liquidity to over 1,400 members to support mortgage lending, economic development and affordable housing in the communities they serve. The Bank is one of 11 Federal Home Loan Banks across the United States and is regulated by the Federal Housing Finance Agency. The Bank is committed to operating in a safe and sound manner and providing value to its members through a variety of products and services.
Learn more about Federal Home Loan Bank of Des Moines
Size
200 employees
Industry
5 Year Trend
+2%
Revenue
$1 billion

Similar Jobs

More Jobs at Federal Home Loan Bank of Des Moines

More Information Technology Jobs

Find similar Manager Site Reliability Engineering jobs: