Senior Vice President, Site Reliability Engineer

BNY Mellon

$150K — $180K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 9+ years in Site Reliability Engineering or Software Engineering
  • Strong programming skills in Java or modern programming languages
  • Experience with observability tools like AppDynamics, Dynatrace, Grafana, or Splunk
  • Hands-on troubleshooting of production systems
  • Strong analytical and problem-solving abilities
  • Proven track record in identifying inefficiencies and implementing automation

Responsibilities

  • Design and implement end-to-end observability across distributed systems
  • Integrate and optimize monitoring tools for better insights
  • Develop real-time visibility through dashboards and alerts
  • Automate repetitive operational tasks and procedures
  • Build self-healing solutions to enhance system reliability
  • Troubleshoot complex production issues effectively
  • Define service health metrics and improve system performance

Benefits

  • Flexible global resources for personal and professional growth
  • Wellbeing programs focused on health and resilience
  • Generous paid leave policies, including volunteer time off
  • Strong culture of excellence with pay-for-performance philosophy
  • Access to tools for achieving financial goals and support during significant life moments
Full Job Description
Job Description

We're seeking a future team member for the role of SVP. Site Reliability Engineer to join our Technology team. This role is located in Lake Mary, FL and Pittsburgh, PA.

In this role, you'll make an impact in the following ways:
• Design and implement end-to-end observability (logs, metrics, traces) across distributed systemsBuild Observability & Monitoring
• Integrate and optimize tools such as AppDynamics, Dynatrace, Grafana, and Splunk
• Develop dashboards, alerts, and telemetry frameworks to provide real-time visibility
• Identify gaps in monitoring and drive adoption of best practices

Drive Automation & Reduce Toil
• Identify repetitive operational work and automate it using code and tooling
• Build self-healing and auto-remediation solutions
• Enable scalable, reliable processes through automation and engineering rigor
• Improve operational efficiency across production environments

Support Production & Incident Triage
• Troubleshoot and resolve complex production issues across distributed systems
• Participate in incident management, triage, and root cause analysis
• Improve monitoring and automation based on recurring incident patterns
• Collaborate with support and engineering teams to improve system stability

Improve Reliability & Performance
• Define and measure service health using SLIs/SLOs and key performance metrics
• Identify system bottlenecks and reliability risks
• Contribute to performance optimization and capacity planning
• Provide input into system architecture to improve resilience and scalability

To be successful in this role, we're seeking the following:
• 9+ years of experience in Site Reliability Engineering, Software Engineering
• Strong programming background in Java (preferred) or another modern language
• Experience with at least one observability platform:
• AppDynamics, Dynatrace, Grafana, or Splunk
• Hands-on experience supporting and troubleshooting production systems
• Strong analytical and problem-solving skills
• Ability to identify inefficiencies and drive automation

Preferred Qualifications
• Experience with distributed systems or microservices architectures
• Familiarity with CI/CD pipelines and DevOps practices
• Exposure to cloud platforms and/or Kubernetes
• Experience scripting (Python, Bash, etc.) for automation
• Knowledge of SRE concepts like observability, incident management, and reliability engineering

Our Benefits and Rewards:

BNY offers highly competitive compensation, benefits, and wellbeing programs rooted in a strong culture of excellence and our pay-for-performance philosophy. We provide access to flexible global resources and tools for your life's journey. Focus on your health, foster your personal resilience, and reach your financial goals as a valued member of our team, along with generous paid leaves, including paid volunteer time, that can support you and your family through moments that matter.

Similar Jobs

More Jobs at BNY Mellon

More Information Technology Jobs

Find similar Senior Vice President, Site Reliability Engineer jobs: