Abbott

Senior Site Reliability Engineer

Abbott$90K — $180K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's in Computer Science, Software Engineering, Systems Engineering, or related discipline; equivalent experience accepted.
  • Minimum 7 years in site reliability, software engineering, or related fields.
  • Excellent communication skills for cross-functional teamwork and technical translation.
  • Strong analytical and debugging abilities under pressure, particularly with production incidents impacting patient safety.
  • Proficiency in system or automation programming languages like Python, Go, Bash, or PowerShell.
  • Demonstrated expertise with Microsoft Azure services relevant to site reliability.
  • Hands-on experience with Kubernetes and Docker for container orchestration.

Responsibilities

  • Design and maintain fault-tolerant, high-availability systems with strict uptime requirements.
  • Identify and resolve performance bottlenecks to ensure real-time responsiveness of services.
  • Define and uphold SLIs, SLOs, and error budgets as part of operational excellence.
  • Develop monitoring and alerting solutions for comprehensive system health insights.
  • Automate operational tasks related to deployment, testing, and recovery processes.
  • Collaborate with cross-disciplinary teams to enforce SRE best practices in sensitive patient data handling.
  • Lead postmortem meetings after incidents to drive actionable improvements.

Benefits

  • Opportunities for professional development and certifications in relevant fields.
  • Access to cutting-edge technologies and a multidisciplinary collaborative environment.
  • Potential to contribute to impactful patient care through SRE practices.
  • Flexibility of working on-site in multiple California locations.
Full Job Description
JOB DESCRIPTION:

About the Role

This Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.

We are seeking a highly skilled and mission-driven Senior Site Reliability Engineer (SRE) to join our DevOps team. In this critical role, you will be responsible for ensuring the reliability, scalability, performance, and operational excellence of Merlin.net - a remote monitoring platform designed to help doctors, cardiologists, and care teams automatically collect and review data from patients with implanted cardiac devices.

This isn't just about keeping servers up; it's about building and maintaining the resilient backbone for systems where failure is not an option, and where our success directly impacts patient care around the world. You will embed within our DevOps team, acting as a bridge between development and operations.

What You'll Do

Design, implement, and maintain highly available, fault-tolerant, and resilient systems that meet demanding uptime and safety requirements. Identify and eliminate performance bottlenecks in software and infrastructure, ensuring low-latency, high-throughput, and real-time responsiveness for customer-facing services. Define, monitor, and uphold Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets. Develop and implement comprehensive monitoring, logging, tracing, and alerting solutions to provide deep insights into system health and behavior at scale. Automate away manual operational tasks, from provisioning and deployment to testing and recovery. Develop and implement strategies for scaling our services and infrastructure to meet evolving business demands, including distributed systems and cloud deployments in Azure. Work closely with software engineering, security, quality, and compliance teams to integrate SRE best practices into our operational processes and infrastructure, ensuring the integrity, availability, and confidentiality of our systems that carry sensitive patient health data. Create clear, concise, and comprehensive documentation, runbooks, and playbooks for operational procedures. Lead blameless postmortem processes following incidents and drive systematic follow-through on action items. Work with a multi-disciplinary team on challenging problems in a fast-paced environment, contributing across architecture reviews, incident response, capacity planning, and reliability roadmap planning.

Required Qualifications
  • Bachelor's in Computer Science, Software Engineering, Systems Engineering, or a related technical discipline; equivalent professional experience will be considered.
  • Minimum 7 years of experience working in site reliability, software Engineering , Systems Engineering or a related technical discipline
  • Excellent communication skills with the demonstrated ability to work effectively in cross-functional teams, translate technical complexity for non-technical stakeholders, and collaborate with development, quality, security, marketing, and regulatory teams.
  • Strong analytical, problem-solving, and debugging skills with a methodical and structured approach to diagnosing complex, distributed system issues under pressure - including production incidents with patient safety implications.
  • Proficiency in at least one systems or automation programming language (e.g., Python, Go, Bash, PowerShell) for building tooling, automation, and operational systems.
  • Demonstrated expertise with Microsoft Azure - including Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, and related managed services.
  • Container orchestration expertise - hands-on production experience with Kubernetes and Docker at scale, including deployment strategies, resource management, and cluster operations.
  • Observability platform experience with tools such as Prometheus, Grafana, the ELK/EFK stack, Datadog, Azure Monitor, or similar enterprise-grade monitoring and tracing platforms.
  • Experience designing and operating CI/CD pipelines for continuous delivery of software in production environments, including safe deployment strategies such as blue/green, canary, and feature flag-gated rollouts.
  • Deep understanding of distributed systems - including load balancing, service meshes, microservices, message queues, and fault-tolerant design patterns.
  • Solid Linux & networking fundamentals - DNS, TCP/IP, HTTP/S, TLS, load balancing, and networking in cloud environments.
  • Incident management experience - including on-call rotation participation, structured incident response, root cause analysis (RCA), and systematic prevention of recurrence.


Preferred Qualifications
  • Experience working in a regulated healthcare, medical device, or life sciences environment, with familiarity in compliance frameworks such as HIPAA.
  • Relevant professional certifications such as: Microsoft Certified Azure DevOps Engineer Expert, Azure Solutions Architect Expert, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Security Specialist (CKS).
  • Background in cost optimization for cloud-native architectures, including FinOps practices for Azure environments.

Experience contributing to or driving disaster recovery (DR) design, business continuity planning, and tabletop exercises.

The base pay for this position is
$90,000.00 - $180,000.00
In specific locations, the pay range may vary from the range posted.

JOB FAMILY:
Product Development

DIVISION:
CRM Cardiac Rhythm Management

LOCATION:
United States > Sunnyvale : 645-647 Almanor Ave

ADDITIONAL LOCATIONS:
United States > Sylmar : 15900 Valley View Court

WORK SHIFT:
Standard

TRAVEL:
Yes, 10 % of the Time

MEDICAL SURVEILLANCE:
No

SIGNIFICANT WORK ACTIVITIES:
Continuous sitting for prolonged periods (more than 2 consecutive hours in an 8 hour day)

About Abbott

Abbott Careers

Joining Abbott means becoming part of a globally diverse team dedicated to making a lasting impact on human health. As a leader in healthcare innovation, Abbott provides a dynamic workplace where careers flourish through growth, leadership, and diversity training.

Opportunities at Abbott

Explore a world of opportunities with our team. Whether you're seeking job opportunities in engineering, marketing, research, or healthcare, Abbott offers a variety of positions that allow professionals to grow their careers. Our commitment to diversity and innovation is evident in every aspect of our work, fostering an inclusive culture that values each team member's contribution.

Work You'll Do

At Abbott, every role contributes to our mission of helping people live fuller lives through better health. From groundbreaking research in medical devices to advancements in pharmaceuticals, our team is at the forefront of healthcare innovation. By joining Abbott, you are not just accepting a job; you are embarking on a path of professional and personal growth.

Internship Programs

Kickstart your career with an Abbott internship. Our programs provide invaluable industry experience and a chance to develop essential skills in a real-world setting. Interns at Abbott work on projects that matter, gaining the experience and knowledge necessary to succeed in their future careers.

Professional Development

Abbott is dedicated to the continuous professional development of its employees. With access to cutting-edge technology, leadership programs, and diversity training, our team members are equipped to lead and innovate within the healthcare industry. We support your career journey with robust training programs, mentorship, and opportunities for networking and professional growth.

Benefits and Culture

Our employees enjoy comprehensive benefits designed to support their life and well-being. From health insurance to retirement plans, we ensure our team has everything they need to thrive. Abbott's culture is built on a foundation of respect and integrity, united by a shared commitment to improving health outcomes.

Join Our Team

Discover the impact you can make with a career at Abbott. We are hiring individuals who are passionate, curious, and driven to lead. Search open positions that match your skills and interests on our Jobs page. Prepare your resume, sharpen your interview skills, and get ready to join a team that's at the cutting edge of healthcare solutions.

Stay Connected

Keep up to date with career tips, industry insights, and company news—all from the people who work here. Subscribe to our Careers Blog and personalize your subscription to receive job alerts and insider tips tailored to your preferences.

Explore Abbott

With a commitment to improving life through innovation, leadership, and diversity, Abbott is a place where you can fulfill your potential. See what exciting and rewarding opportunities await at Abbott by exploring our career opportunities today.

SEARCH ABBOTT JOBS

Join us in our mission to make the world a healthier place through innovation, leadership, and diversity. Your journey to a fulfilling career at Abbott starts here.
Learn more about Abbott
Size
113,000 employees
Market Cap
$189 billion
Industry
Net Income
$4.4 billion
Founded
1944
5 Year Trend
+15.6%
Revenue
$34.6 billion
NASDAQ

Similar Jobs

More Jobs at Abbott

More Information Technology Jobs

Find similar Senior Site Reliability Engineer jobs: