JP Morgan Chase & Co.

Lead Site Reliability Engineer

JP Morgan Chase & Co.$130K — $155K *
Plano, TX 75024In-Person
Enterprise Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years experience in site reliability engineering or related domain.
  • Proficiency in a programming language (Python, Java/Spring Boot, .Net).
  • Strong foundation in reliability, scalability, and security best practices.
  • Demonstrated experience leveraging enterprise AI capabilities for SRE workflows.
  • Experience with observability techniques like monitoring and telemetry collection.

Responsibilities

  • Champion site reliability culture and document best practices.
  • Lead initiatives to enhance application reliability using data analytics.
  • Facilitate collaboration on service level objectives and error budgets.
  • Serve as primary point of contact during major incidents.
  • Mentor team members and provide technical expertise in SRE practices.
  • Drive adoption of AI-assisted reliability workflows across SDLC processes.

Benefits

  • Leadership role in a prestigious global firm.
  • Direct impact on technology and business outcomes.
  • Access to advanced AI capabilities for operational efficiency.
  • Opportunities for skills development and professional growth.
  • Collaborative work environment focused on innovation.
Full Job Description
JOB DESCRIPTION

Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.

As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team , you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them. Take lead and conduct resiliency design reviews, break up complex problems into digestible work for other engineers, act as a technical lead for medium to large-sized products, and provide advice and mentoring to other engineers.

Job Responsibilities
  • Consistently models and champions site reliability culture and practices, documents and shares knowledge within your organization via internal forums and communities of practice
  • Leads initiatives to improve the reliability and stability of your team27s applications and platforms using data-driven analytics to improve service levels, proactively identifying and solving technology-related bottlenecks in areas of expertise
  • Drives collaboration with your team to identify comprehensive service level indicators and the stakeholder partners to establish reasonable service level objectives and error budgets with your customers
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Serves as the main point of contact during major incidents for your application and has the skills to identify and solve the issue quickly to avoid financial loss to the business
  • Offers a high level of technical expertise within one or more technical domains and provides advice and mentorship to other engineers
  • Leads reuse-first adoption of AI-assisted reliability workflows across SDLC/toolchain practices (e.g., CI/CD quality checks, test/validation automation, and operational readiness), ensuring traceability/auditability, resiliency, and security controls.

    Required qualifications, capabilities, and skills
  • A. Formal training or certification on site reliability engineering concepts and 5+ years applied experience ( NAMR/APAC 2D India/ LATAM/ Hong Kong)
    B. Formal training or certification on site reliability engineering concepts and advanced applied experience (EMEA/LATAM-Brazil)
    C. Singapore follow local country guidance
  • Demonstrated proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices
  • Fluent in at least one programming language such as: Python, Java/Spring Boot, .Net
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity.
  • Ability to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails for team usage, and ensure outcomes align to resiliency and security expectations.
  • Proficient knowledge and experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection
  • Proficient with continuous integration and continuous delivery practices and tooling
  • Proficient with container and container orchestration
  • Experience with troubleshooting common networking technologies and issues
  • Advanced knowledge of software applications and technical processes with emerging depth in one or more technical disciplines, and actively self-educates to evaluate and recommend suitable new technologies

    Preferred qualifications, capabilities, and skills
  • Experience implementing and managing SLOs/SLIs, error budgets, and operational readiness reviews for distributed systems, including leading post-incident analysis and resilience improvements.
  • Hands-on expertise in observability and monitoring tools including Grafana, Dynatrace, Prometheus, Datadog, and Splunk; experience with SLO alerting, white/black box monitoring, and telemetry collection.
  • Advanced proficiency in Python and/or Java for building automation, tooling, and operational workflows.
  • Experience with CI/CD tools and practices including Jenkins, GitLab, and Terraform; strong grasp of DevOps principles and continuous delivery pipelines.
  • Strong incident management experience; effective under pressure with excellent stakeholder communication and the ability to drive root-cause analysis and auto-remediation.
  • Familiarity with containerization and orchestration technologies such as Docker, Kubernetes, and ECS; hands-on experience with infrastructure-as-code is a plus.
  • Exposure to public cloud platforms (AWS or equivalent), including compute, storage, networking, messaging, and infrastructure automation (CloudFormation, Terraform) is desirable.

About JP Morgan Chase & Co.

JP Morgan Chase & Co. stands at the forefront of the global financial services industry. They offer an expansive array of products and services to a diverse clientele, including individuals, corporations, governments, and institutions. Ever since the merger of J.P. Morgan & Co. and Chase Manhattan Corporation in 2000, this industry-leading entity has become renowned for its comprehensive portfolio encompassing consumer and community banking, corporate and investment banking, commercial banking, as well as asset and wealth management. Headquartered in the vibrant city of New York, JP Morgan Chase & Co. boasts a formidable presence across over 100 countries worldwide.

Unveiling Employment Opportunities at JP Morgan Chase & Co.

Vacancies and Hiring Initiatives

JP Morgan Chase & Co. is continuously on the lookout for talented individuals eager to contribute to its legacy of excellence. The company's recruitment efforts are geared towards identifying candidates with the right blend of skills and qualifications to drive forward its various business segments. Whether you are a seasoned professional or a recent graduate, JP Morgan Chase offers a plethora of job openings across multiple disciplines.

High-Demand Positions

Among the myriad of roles, certain positions stand out for their attractive compensation packages and career advancement prospects. Notably, high-paying jobs at JP Morgan Chase & Co. include Relationship Manager, Branch Manager, and Software Engineer. These roles are critical to the firm's operations and offer lucrative opportunities for those with the requisite expertise.

Navigating the Job Market at JP Morgan Chase & Co.

Leveraging Job Portals and Job Alerts

For job seekers aiming to tap into the opportunities at JP Morgan Chase, staying updated through job portals and subscribing to job alerts is crucial. These tools can provide timely information about job openings, job fairs, and recruitment events, enabling candidates to apply promptly and prepare adequately for interviews.

Preparing Your Job Application

Your job application, comprising your resume and cover letter, is your ticket to securing an interview at JP Morgan Chase. Highlight your qualifications, skills, and experiences that align with the job listing, ensuring you stand out in the competitive job market.

Acing the Interview

Preparation is key to succeeding in your interview with JP Morgan Chase. Familiarize yourself with the company's business segments, values, and recent achievements. Demonstrating how your background and aspirations match the company's goals can significantly increase your chances of employment. A World of Job Opportunites in the Financial Services Industry JP Morgan Chase & Co. offers a world of job opportunities for those seeking to make their mark in the financial services industry. With competitive salaries, comprehensive benefits, and endless possibilities for growth, positions at JP Morgan Chase are highly coveted. By staying informed through job sites, tailoring your applications, and preparing thoroughly for interviews, you can enhance your prospects of joining the esteemed ranks of JP Morgan Chase employees. Explore the job board, seize the job opportunities, and embark on a rewarding career journey with one of the world's leading financial institutions.
Learn more about JP Morgan Chase & Co.
Size
661 employees
Market Cap
$384.5 billion
Industry
Net Income
$29.1 billion
Founded
1823
5 Year Trend
+0.7%
Revenue
$261.5 million
NASDAQ

Similar Jobs

More Jobs at JP Morgan Chase & Co.

More Enterprise Technology Jobs

Find similar Lead Site Reliability Engineer jobs: