PagerDuty

Site Reliability Engineer I

PagerDuty$98K — $148K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 0 to 1+ years in Site Reliability Engineering, DevOps, or Platform Engineering roles
  • Experience with Linux-based systems in production
  • Knowledge of networking fundamentals like load balancing and DNS
  • Familiarity with container orchestration tools such as Kubernetes
  • Experience with cloud-native infrastructure (e.g., AWS, GCP, Azure)
  • Proficiency in at least one programming language (e.g., Python, Ruby, Go)
  • Experience with Infrastructure as Code tools (e.g., Terraform, CloudFormation)

Responsibilities

  • Support foundational infrastructure, including Kubernetes clusters and networking systems
  • Contribute to the reliability and scalability of PagerDuty's core platform
  • Participate in agile rituals and communicate progress early
  • Stay updated on technical trends and propose innovative solutions
  • Monitor system health and participate in 24/7 on-call incident response

Benefits

  • Comprehensive benefits package
  • Flexible work arrangements
  • Company equity options
  • Generous vacation time and paid holidays
  • Paid parental leave of up to 22 weeks
  • 20 hours of paid volunteer time off annually
  • Company-wide wellness initiatives
Full Job Description
As a Site Reliability Engineer I on the Core Infrastructure team in our Atlanta office, you'll help build and operate the foundational infrastructure that powers PagerDuty's real-time digital operations platform. Our systems support millions of events and alerts daily, enabling customers to detect, respond to, and resolve incidents quickly and reliably. You'll work at the intersection of platform evolution and operational excellence, building and evolving foundational network, compute, and ingress infrastructure while scaling and hardening existing systems. Your work will directly impact the reliability, scalability, and security of the services our customers rely on to keep their businesses running as PagerDuty continues to grow across products, regions, and customer use cases. Key Responsibilities • Support and improve foundational infrastructure, including networking, compute platforms, Kubernetes clusters, and ingress/traffic management systems. • Contribute to the reliability and scalability of PagerDuty's core platform by hardening existing systems and supporting the rollout of new infrastructure capabilities. • Participate in agile rituals (standups, planning, retros) and communicate progress/risks early • You stay current on technical trends to suggest innovative tools and approaches to interesting problems • Monitor system health using metrics, logs, and alerts, and participate in 24/7 on-call rotations to help detect, respond to, and resolve incidents. Basic Qualifications • 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles • Hands-on experience operating Linux-based systems in production environments • Working knowledge of networking fundamentals, such as load balancing, DNS, TLS, and ingress traffic flow • Experience with container orchestration (e.g., EKS, Kubernetes) • Experience working on cloud-native infrastructure (e.g., AWS, GCP, Azure), including networking and compute concepts • Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.) • Experience with Infrastructure as Code (e.g., Terraform, CloudFormation) Preferred Qualifications • Experience with AWS cloud networking concepts such as VPCs, subnets, routing, security groups, and load balancers • Experience operating or contributing to production Kubernetes platforms (e.g., EKS), including cluster upgrades, networking, or ingress configuration • Experience with monitoring, observability, and logging platforms (e.g., DataDog, New Relic, SumoLogic, Splunk, Prometheus, Grafana) • Familiarity with service meshes, ingress controllers, or API gateways (e.g., Envoy, Istio, NGINX) Salary Range: $98,000 to $148,500 Hesitant to apply? We encourage you to submit your resume even if you don't meet every requirement. We value potential and consider each candidate's full professional story. Whether you're exploring a career change or taking your next step, we look forward to reviewing your application. If this just isn't the right role or time - sign up for job alerts! Where we work PagerDuty operates a hybrid work model with offices in 8 major cities: Atlanta, Lisbon, London, San Francisco, Santiago, Sydney, Tokyo, and Toronto. While we offer flexibility within our established locations, we cannot employ candidates residing in: Location restrictions: Australia: Northern Territory, Queensland, South Australia, Tasmania, Western Australia Canada: Alberta, Manitoba, Newfoundland, Northwest Territories, Nunavut, PEI, Quebec, Saskatchewan, Yukon United States: Alaska, Hawaii, Iowa, Louisiana, Mississippi, Nebraska, New Mexico, Oklahoma, Rhode Island, South Dakota, West Virginia, Wyoming Candidates must reside in an eligible location, which vary by role. How we work Our values guide how we support customers, collaborate with colleagues, develop products, and foster a culture of belonging. They define not just our actions, but what it means to be Dutonian. People Leaders at PagerDuty are responsible for creating high performance environments that drive accountability. PagerDuty has four key dimensions that define our Leadership Impact: Lead Self, Lead the Team, Lead the Business, and Lead the Future. Each dimension has three associated competencies to give leaders a shared language for guiding their development, career, promotion, and succession planning discussions. Our Manager Expectations serve as a practical guide for managers to understand their responsibilities, prioritize their efforts, and drive engagement and performance. What we offer As a global organization, our total rewards approach is competitive with industry standards and aligned with local laws and regulations. Learn more, including country-specific offerings, on our benefits site. Your package may include: • Competitive salary • Comprehensive benefits package • Flexible work arrangements • Company equity* • ESPP (Employee Stock Purchase Program)* • Retirement or pension plan* • Generous paid vacation time • Paid holidays and sick leave • Dutonian Wellness Days & HibernationDuty - companywide paid days off in addition to PTO • Paid parental leave: 22 weeks for pregnant parent, 12 weeks for non-pregnant parent (some countries have longer leave standards and we comply with local laws)* • Paid volunteer time off: 20 hours per year • Company-wide hack weeks • Mental wellness programs *Eligibility may vary by role, region, and tenure

About PagerDuty

PagerDuty is a digital operations management platform that provides reliable notifications, automatic escalations, on-call scheduling, and other functionality to help teams detect and fix infrastructure problems quickly. The company's platform integrates with over 350 third-party tools and services, including AWS, Slack, and Zoom, and is used by over 14,000 organizations worldwide. PagerDuty was founded in 2009 and is headquartered in San Francisco, California.
Learn more about PagerDuty
Size
950 employees
Market Cap
$2.3 billion
Industry
Net Income
-$57.2 million
Founded
2009
Revenue
$200.2 million
NASDAQ

Similar Jobs

More Jobs at PagerDuty

More Information Technology Jobs

Find similar Site Reliability Engineer I jobs: