JP Morgan Chase & Co.

Senior Lead Site Reliability Engineer

JP Morgan Chase & Co.$145K — $175K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Formal training or certification in software engineering concepts with 5+ years of applied experience.
  • Experience in Site Reliability Engineering (SRE), DevOps, or production engineering roles.
  • Hands-on expertise in managing Kubernetes workloads for deployments and debugging.
  • Practical experience with AWS services (EKS, ECS, Lambda, DynamoDB, S3) in production environments.
  • Competency in CI/CD tools like Spinnaker and/or Harness for reliable deployments.
  • Proficiency in Infrastructure as Code (IaC) using Terraform, along with scripting skills in Python, Bash, or Go.
  • Strong skills in incident response, root cause analysis (RCA), and remediation activities.

Responsibilities

  • Own production reliability outcomes by overseeing operational health metrics.
  • Define and evolve service level indicators and objectives, focusing on actionable alerting.
  • Enhance observability and troubleshooting across Kubernetes and AWS environments.
  • Lead incident response efforts, driving triage and recovery processes effectively.
  • Manage Kubernetes workloads, ensuring optimal scaling and resilience patterns.
  • Operate AWS container and serverless technologies with an emphasis on scaling and safety.
  • Improve release engineering processes to boost deployment safety using modern tooling.

Benefits

  • Opportunity to work in a rapidly growing technology field.
  • Access to cutting-edge AWS and Kubernetes technologies.
  • Involvement in significant decision-making impacting the company’s architecture.
  • Opportunity for impactful work that shapes target state architecture.
  • Professional development through knowledge-sharing within a dynamic team environment.
Full Job Description
JOB DESCRIPTION
There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. 

As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Commercial Investment Banking team of Fraud Prevention, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform. 

 You are an integral part of a team that works to develop high-quality architecture solutions for various software applications and platform products. You drive significant business impact and help shape the target state architecture through your capabilities in multiple architecture domains. You will ensure the platform is reliable, secure, performant, and resilient in production across Kubernetes-based environments and AWS. You will apply SRE principles to drive measurable improvements in availability and latency, reduce operational toil through automation, and strengthen deployment safety and recovery capabilities in close partnership with engineering and platform teams.


Job responsibilities

  • Own production reliability outcomes by managing day-to-day operational health (availability, latency, throughput, error rates), proactively surfacing risks, and driving remediation.
  • Define and evolve service level indicators/service level objectives (SLIs/SLOs) and error budgets; build actionable, customer-impact-aligned alerting and reduce noise through tuning and standardization.
  • Improve end-to-end observability and troubleshooting (metrics, logs, traces), dashboards, and runbooks across Kubernetes and Amazon Web Services (AWS); perform deep technical triage of distributed-system issues.
  • Lead incident response and problem management by participating in on-call, driving triage/mitigation/recovery, completing root cause analyses (RCAs), and ensuring corrective and preventive actions close.
  • Operate Kubernetes workloads including autoscaling, rollout/rollback procedures, resource tuning, and resilience patterns for containerized services.
  • Operate AWS container and serverless components (for example, Amazon Elastic Kubernetes Service/Elastic Container Service/AWS Lambda) with a focus on scaling, retries, and safe failure modes.
  • Improve release engineering and delivery reliability by increasing the safety and repeatability of deployments using Spinnaker and Harness.
  • Build infrastructure as code and environment consistency by developing and maintaining Terraform modules and automation for reliable, repeatable environments.
  • Strengthen database and data-service reliability by partnering with engineering and platform teams to improve reliability patterns across multiple database technologies and data services (for example, DynamoDB, Amazon Simple Storage Service).
  • Embed security and controls into operations by applying secure operational practices and ensuring processes meet required control standards.
  • Lead small-to-medium initiatives end-to-end from proposal through production adoption, using enterprise-authorized AI capabilities to accelerate triage and toil reduction while validating outputs and handling operational data per sensitivity and security requirements.
Required qualifications, capabilities, and skills
  • Formal training or certification on software engineering concepts and 5+ years applied experience
  • Experience in SRE/DevOps/production engineering or equivalent
  • Hands-on experience operating Kubernetes workloads (deployments, scaling, debugging)
  • Practical experience with AWS (EKS, ECS, Lambda, Dynamo DB, S3) in production
  • Experience with CI/CD and release tooling such as Spinnaker and/or Harness
  • Proficiency with Terraform (IaC), and scripting/automation (Python/Bash/Go)
  • Strong incident response skills, RCA writing, and ability to drive remediation work
  • Solid fundamentals in Linux, networking, and troubleshooting distributed systems
  • Ability to independently execute well-scoped reliability work and escalate when needed
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements
 
Preferred qualifications, capabilities, and skills
  • Experience implementing SLO programs and alerting aligned to customer journeys
  • Experience with performance testing, capacity planning, and resilience testing (fault injection/chaos, DR exercises)
  • Experience improving operational maturity: standardized runbooks, automated health checks, auto-remediation, and deployment guardrails
  • Experience with fraud screening/decisioning or payment flows
  • Familiarity with database reliability patterns (capacity, backups, failover readiness)
  • Experience with secure operational practices (least privilege, secrets handling)
  • Experience partnering with engineering and platform teams to drive reliability improvements

About JP Morgan Chase & Co.

JP Morgan Chase & Co. stands at the forefront of the global financial services industry. They offer an expansive array of products and services to a diverse clientele, including individuals, corporations, governments, and institutions. Ever since the merger of J.P. Morgan & Co. and Chase Manhattan Corporation in 2000, this industry-leading entity has become renowned for its comprehensive portfolio encompassing consumer and community banking, corporate and investment banking, commercial banking, as well as asset and wealth management. Headquartered in the vibrant city of New York, JP Morgan Chase & Co. boasts a formidable presence across over 100 countries worldwide.

Unveiling Employment Opportunities at JP Morgan Chase & Co.

Vacancies and Hiring Initiatives

JP Morgan Chase & Co. is continuously on the lookout for talented individuals eager to contribute to its legacy of excellence. The company's recruitment efforts are geared towards identifying candidates with the right blend of skills and qualifications to drive forward its various business segments. Whether you are a seasoned professional or a recent graduate, JP Morgan Chase offers a plethora of job openings across multiple disciplines.

High-Demand Positions

Among the myriad of roles, certain positions stand out for their attractive compensation packages and career advancement prospects. Notably, high-paying jobs at JP Morgan Chase & Co. include Relationship Manager, Branch Manager, and Software Engineer. These roles are critical to the firm's operations and offer lucrative opportunities for those with the requisite expertise.

Navigating the Job Market at JP Morgan Chase & Co.

Leveraging Job Portals and Job Alerts

For job seekers aiming to tap into the opportunities at JP Morgan Chase, staying updated through job portals and subscribing to job alerts is crucial. These tools can provide timely information about job openings, job fairs, and recruitment events, enabling candidates to apply promptly and prepare adequately for interviews.

Preparing Your Job Application

Your job application, comprising your resume and cover letter, is your ticket to securing an interview at JP Morgan Chase. Highlight your qualifications, skills, and experiences that align with the job listing, ensuring you stand out in the competitive job market.

Acing the Interview

Preparation is key to succeeding in your interview with JP Morgan Chase. Familiarize yourself with the company's business segments, values, and recent achievements. Demonstrating how your background and aspirations match the company's goals can significantly increase your chances of employment. A World of Job Opportunites in the Financial Services Industry JP Morgan Chase & Co. offers a world of job opportunities for those seeking to make their mark in the financial services industry. With competitive salaries, comprehensive benefits, and endless possibilities for growth, positions at JP Morgan Chase are highly coveted. By staying informed through job sites, tailoring your applications, and preparing thoroughly for interviews, you can enhance your prospects of joining the esteemed ranks of JP Morgan Chase employees. Explore the job board, seize the job opportunities, and embark on a rewarding career journey with one of the world's leading financial institutions.
Learn more about JP Morgan Chase & Co.
Size
661 employees
Market Cap
$384.5 billion
Industry
Net Income
$29.1 billion
Founded
1823
5 Year Trend
+0.7%
Revenue
$261.5 million
NASDAQ

Similar Jobs

More Jobs at JP Morgan Chase & Co.

More Information Technology Jobs

Find similar Senior Lead Site Reliability Engineer jobs: