ServiceNow

Senior Software Engineer - SRE & AIOps

ServiceNow$143K — $243K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years in software engineering or infrastructure operations, with 3+ years in SRE, DevOps, or cloud platform roles; or equivalent work experience.
  • 2+ years of hands-on experience with production Kubernetes clusters.
  • Proficient in at least one Infrastructure-as-Code tool (e.g., Terraform, CloudFormation).
  • Demonstrable experience with major cloud platforms (AWS, Azure, or GCP).
  • Strong foundation in Linux system administration and scripting (Python, Go, or Bash).

Responsibilities

  • Deploy, operate, and troubleshoot Kubernetes clusters in hybrid and multi-cloud environments.
  • Implement closed-loop auto-remediation systems to resolve infrastructure failures.
  • Design and evolve SRE tooling stack for monitoring and incident management.
  • Develop SLO frameworks and alerting policies for on-call engineers.
  • Build Infrastructure-as-Code frameworks and GitOps pipelines for reproducible deployments.

Benefits

  • Comprehensive health plans including flexible spending accounts.
  • 401(k) Plan with company match.
  • Employee Stock Purchase Plan (ESPP).
  • Matching donations program.
  • Flexible time away plans and family leave programs.
Full Job Description
Job Description

About the role:

ServiceNow is seeking a Senior Software Engineer - SRE & AIOps to contribute to infrastructure automation, operational resilience, and toil elimination across our hybrid cloud and data center operations. Embedded within the Site Reliability & Database Engineering organization, you will implement automation-first systems that reduce manual intervention, accelerate incident remediation, and enable our global engineering teams to operate reliably at scale.

This role combines solid hands-on technical expertise in Kubernetes, cloud platforms, and DevOps practices with growing technical leadership capabilities. You will contribute to SRE tooling design, develop auto-remediation capabilities, and help establish patterns that maintain ServiceNow's cloud platform reliability while minimizing operational toil across follow-the-sun global teams.

What you get to do in this role:
  • Deploy, operate, and troubleshoot production Kubernetes clusters across hybrid and multi-cloud environments, maintaining operational standards and supporting high-velocity application deployments.
  • Implement and maintain closed-loop auto-remediation systems that detect, classify, and resolve transient infrastructure failures, leveraging automation frameworks and machine learning insights to reduce MTTR and on-call burden.
  • Contribute to the design and evolution of SRE tooling stack, including monitoring platforms, incident management systems, log aggregation, and observability integrations that support global on-call operations.
  • Develop and maintain SLO frameworks, alerting policies, and automated runbooks that empower on-call engineers to resolve issues autonomously while managing alert fatigue.
  • Build and maintain Infrastructure-as-Code frameworks and GitOps pipelines that enable reproducible infrastructure deployments across hybrid and multi-cloud environments with security and compliance guardrails.
  • Support hybrid cloud and data center operations, including on-premises infrastructure, public cloud environments, and workload optimization across multi-region deployments.
  • Contribute to adoption of containerization, microservices, and DevOps patterns across engineering teams, establishing CI/CD best practices and network security controls.
  • Support on-call rotation operations and incident response processes across different time zones, helping develop runbooks and contributing to post-incident reviews that drive continuous improvement.
  • Share knowledge and mentor junior SRE engineers on reliability patterns, incident investigation techniques, and automation best practices.
  • Champion a culture of blameless incident analysis, data-driven decision-making, and continuous improvement through knowledge sharing and documentation.
  • Identify and systematically automate repetitive operational tasks, from infrastructure provisioning to incident response, improving team efficiency and capacity.


Qualifications

To be successful in this role you have:
  • Kubernetes Proficiency: Solid hands-on experience operating production Kubernetes clusters, including deployment models, pod orchestration, resource management, network policies, and troubleshooting runtime issues.
  • Incident Remediation Experience: Demonstrated experience designing and implementing automated remediation systems, including alert automation, runbook development, and self-healing mechanisms.
  • Cloud Platform Knowledge: Strong hands-on experience with AWS (EKS, EC2, RDS) and/or Azure (AKS, VMs) or GCP (GKE), with understanding of core SRE-related services.
  • DevOps & IaC Skills: Solid experience with Infrastructure-as-Code tools (Terraform, CloudFormation) and GitOps practices.
  • SRE Tooling Familiarity: Working knowledge of observability platforms, incident management systems, and log aggregation tools.
  • Distributed Systems Understanding: Understanding of distributed system challenges, fault tolerance, and resilience patterns.
  • On-Call Operations: Experience participating in on-call rotations and understanding 24/7 operational models, runbook development, and escalation procedures.
  • Cloud & Hybrid Operations: Hands-on experience working with cloud infrastructure and understanding hybrid cloud concepts.
  • Systems Administration: Strong foundation in Linux system administration, performance troubleshooting, and scripting (Python, Go, or Bash).
  • Collaborative Mindset: Ability to work effectively with infrastructure and application teams, contribute to technical discussions, and help drive reliability improvements.

Qualifications
  • Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI's potential impact on the function or industry.
  • 5+ years in software engineering or infrastructure operations, with 3+ years in SRE, DevOps, or cloud platform engineering roles with a Bachelor's degree; or 3 years and a Master's degree; or a PhD without experience; or equivalent work experience.
  • 2+ years of hands-on experience working with production Kubernetes clusters.
  • Proficiency in at least one Infrastructure-as-Code tool: Terraform, CloudFormation, or equivalent.
  • Demonstrable hands-on experience with at least one major cloud platform: AWS, Azure, or GCP.
  • Experience operating in on-call environments and participating in incident response.
  • Experience implementing or improving automated remediation and alert systems.
  • Strong foundation in Linux system administration, performance troubleshooting, and scripting (Python, Go, Bash).
  • Demonstrated commitment to reliability engineering and continuous improvement through hands-on contributions.
  • Bachelor's degree in computer science, Computer Engineering, or related field (or equivalent professional experience).

Preferred:
  • Kubernetes certification (CKA, CKAD, or equivalent).
  • Experience with service mesh technologies or advanced Kubernetes networking.
  • Background in cloud migration or infrastructure modernization projects.
  • Experience with cost optimization in cloud environments.
  • Track record of implementing automation solutions that significantly reduced operational toil.

Why This Role?

This role offers the opportunity to work with infrastructure automation and reliability engineering at scale. You will implement systems and practices that directly reduce operational burden across ServiceNow's global engineering teams. Your contributions will help establish reliable, automated infrastructure operations and provide a clear career path toward senior technical leadership. This is a role for an engineer who enjoys solving complex operational challenges, continuous learning, and working collaboratively to improve how systems operate.

For positions in this location, we offer a base pay of $143,200 - $243,400, plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location. We also offer health plans, including flexible spending accounts, a 401(k) Plan with company match, ESPP, matching donations, a flexible time away plan and family leave programs. Compensation is based on the geographic location in which the role is located and is subject to change based on work location.

Additional Information

Work Personas

We approach our distributed world of work with flexibility and trust. Work personas (flexible, remote, or required in office) are categories that are assigned to ServiceNow employees depending on the nature of their work and their assigned work location. Learn more here. To determine eligibility for a work persona, ServiceNow may confirm the distance between your primary residence and the closest ServiceNow office using a third-party service.

About ServiceNow

ServiceNow provides cloud-based solutions that define, structure, manage, and automate services for enterprise operations in North America, Europe, the Middle East, Africa, the Asia Pacific, and other countries. The company offers service management solutions, including incident, problem, change, request, and cost management as well as service catalogs; and IT, HR, facilities, and field service management solutions. It also provides IT operations management solutions covering service mapping, delivery, and assurance solutions; business management solutions such as financial management, project portfolio suite, vendor performance management, and performance analytics as well as governance, risk, and compliance; and application development services.

ServiceNow Careers

Join the dynamic team at ServiceNow, a global leader in digital workflow solutions, where innovation and leadership converge to shape the future of work. At ServiceNow, we offer more than just job opportunities; we provide a platform for professional growth and a chance to be part of a culture that values diversity, creativity, and continuous learning.

Work You’ll Do

Embark on a career journey with ServiceNow and contribute to the world’s leading enterprises' digital transformation. Our team is at the forefront of developing cutting-edge technologies that improve how people work. With ServiceNow, you will use your skills to impact businesses and industries profoundly, driving efficiency and innovation.

Join Our Market-Leading Team

ServiceNow is not just another technology company. We are a team that thrives on diversity and leadership, fostering an inclusive environment that promotes growth and development. Our commitment to diversity training ensures that every team member can achieve their potential.

Innovative Work

ServiceNow is home to more than 10,000 dedicated professionals who lead the charge in digital workflows and enterprise solutions. As part of our team, you will engage in projects that merge technology with practical applications, creating revolutionary products that advance how services are delivered and managed.

Career Development

At ServiceNow, your career trajectory is filled with boundless opportunities. We support your growth with robust training programs, leadership development courses, and access to global challenges. Whether you are looking for an internship, full-time position, or leadership role, ServiceNow equips you with the tools to excel.

Be Part of a Great Team

Working at ServiceNow means being part of a community that values teamwork and innovation. Our collaborative environment encourages networking and sharing ideas, making our workplace vibrant and dynamic. The benefits of joining ServiceNow extend beyond comprehensive health and wellness; they include fostering professional connections and friendships that last a lifetime.

Explore Job Opportunities and Internships

Whether you’re a seasoned professional or a recent graduate, ServiceNow offers a range of employment options to suit your career goals. From internships that provide real-world experience to full-time positions that challenge you to leverage your expertise, we are committed to hiring the best talent.

Stay Connected

Join Our Team Search open positions that match your skills and interests. At ServiceNow, we look for passionate, curious, and solution-driven team players. Explore the possibilities that await you at a company that is committed to your professional success.

Keep Up to Date

Stay ahead with career tips, insider perspectives, and industry-leading insights you can put to use today—all from the people who work here.

Job Alert Emails

Customize your subscription to receive job alerts, the latest news, and insider tips tailored to your preferences. Discover the exciting and rewarding career opportunities that await at ServiceNow.

ServiceNow Careers

Empowering professionals to achieve more, ServiceNow is where careers are future-proofed, and ambitions are realized. Join us in our journey of growth and innovation.
Learn more about ServiceNow
Size
16,881 employees
Market Cap
$76.5 billion
Industry
Net Income
$118.5 million
Founded
2004
5 Year Trend
+33.5%
Revenue
$4.5 billion
NASDAQ

Similar Jobs

More Jobs at ServiceNow

More Information Technology Jobs

Find similar Senior Software Engineer - SRE & AIOps jobs: