NCR Corporation

Senior Site Reliability Engineer - Unified Observability

NCR Corporation$120K — $150K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, IT, Engineering, or equivalent experience.
  • Over 10 years in Site Reliability Engineering, Cloud Engineering, or related fields.
  • Proven experience in designing enterprise-scale observability platforms.
  • Deep expertise in Kubernetes, specifically AKS and GKE.
  • Strong background in Azure and Google Cloud services.
  • Hands-on with observability tools like Grafana and Datadog.
  • Experience with SLIs, SLOs, and reliability metrics.

Responsibilities

  • Lead architecture and design of enterprise observability solutions.
  • Establish observability standards for monitoring and operational analytics.
  • Develop real-time dashboards for infrastructure and application visibility.
  • Define and implement reliability frameworks and operational KPIs.
  • Collaborate with cross-functional teams to mitigate reliability risks.
  • Improve incident management through automation and best practices.
  • Integrate observability capabilities with existing operational workflows.

Benefits

  • Opportunity to work on a cutting-edge observability platform.
  • Mentorship and leadership development in a key technical role.
  • Cross-functional collaboration with diverse teams.
  • Involvement in enterprise AI-driven observability initiatives.
  • Potential for significant influence in technical strategy.
Full Job Description
Position Overview

We are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer Unified Observability initiative. This strategic role will be responsible for building and evolving a unified enterprise observability platform that delivers end-to-end visibility across NCR Voyix Restaurants, Retail, and Payments environments.

The ideal candidate will bring 10+ years of experience in Site Reliability Engineering, Platform Engineering, Cloud Operations, or related disciplines, with a proven track record of driving enterprise-scale observability, reliability, and operational excellence. This individual must be comfortable operating across organizational boundaries and partnering closely with Product Engineering, Infrastructure, Security, Operations, Architecture, and Executive Leadership teams to establish a comprehensive observability strategy and improve platform resilience.

This role will serve as a key technical leader responsible for defining standards, influencing architecture decisions, and enabling proactive operations through unified monitoring, telemetry, automation, and AI-driven insights.

Key Responsibilities
  • Lead the architecture, design, implementation, and continuous improvement of enterprise observability solutions across Azure, Google Cloud Platform (GCP), Kubernetes, and hybrid environments.
  • Establish and drive enterprise observability standards for monitoring, logging, distributed tracing, telemetry, and operational analytics.
  • Develop and maintain executive, operational, and engineering dashboards that provide real-time visibility into infrastructure, applications, platform health, customer experience, and business transactions.
  • Define, evangelize, and implement reliability frameworks including SLIs, SLOs, error budgets, operational KPIs, and service health metrics.
  • Partner cross-functionally with Engineering, Infrastructure, Security, Product, and Operations teams to identify reliability risks and drive operational excellence initiatives.
  • Lead efforts to improve incident prevention, detection, response, and recovery through intelligent alerting, automation, event correlation, and observability best practices.
  • Integrate observability capabilities with ServiceNow, CI/CD pipelines, automation frameworks, and enterprise operational workflows.
  • Influence technical strategy and roadmap decisions related to reliability engineering, platform observability, and operational readiness.
  • Support and drive enterprise initiatives involving AI-driven observability, predictive analytics, anomaly detection, and event intelligence.
  • Mentor engineers and serve as a subject matter expert for observability, reliability engineering, and cloud-native operations.
  • Establish governance, adoption, and best practices across multiple product and engineering teams to ensure consistent observability standards enterprise-wide.

Required Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.
  • 10+ years of experience in Site Reliability Engineering, Cloud Engineering, Platform Engineering, DevOps, or related technical disciplines.
  • Demonstrated success designing and operating observability platforms in large-scale enterprise environments.
  • Deep expertise with Kubernetes platforms, including AKS and GKE.
  • Strong experience with Azure and Google Cloud Platform services and architectures.
  • Hands-on experience with enterprise observability tools such as Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, or similar platforms.
  • Advanced knowledge of monitoring, logging, telemetry collection, distributed tracing, and observability engineering principles.
  • Experience defining and operationalizing SLIs, SLOs, error budgets, reliability metrics, and service health frameworks.
  • Strong automation and Infrastructure as Code expertise using Terraform and related tools.
  • Proficiency developing automation solutions using Python, Go, PowerShell, or similar languages.
  • Experience integrating observability solutions into CI/CD pipelines and modern DevOps workflows.
  • Proven ability to influence technical direction and collaborate effectively with stakeholders across Engineering, Product, Infrastructure, Security, and Operations organizations.
  • Strong communication, leadership, and stakeholder management skills with the ability to translate technical concepts for both technical and business audiences.

Preferred Qualifications
  • Experience leading enterprise observability transformations or platform modernization initiatives.
  • Experience with AI Ops, event correlation, operational analytics, and predictive monitoring capabilities.
  • Knowledge of ServiceNow integrations and ITSM/ITOM processes.
  • Experience supporting highly available, customer-facing SaaS platforms at scale.
  • One or more cloud certifications (Azure, Google Cloud, Kubernetes, or related technologies).
  • Previous experience serving as a technical lead, mentor, or architect within a reliability engineering organization.


Offers of employment are conditional upon passage of screening criteria applicable to the job

About NCR Corporation

Radiant Systems, Inc. (Nasdaq: RADS) is a global provider of innovative technology to the hospitality and retail industries. For more than two decades, Radiant's point of sale hardware and software solutions have helped to redefine the consumer experience in more than 100,000 restaurants, retail stores, stadiums, parks, arenas, cinemas, convenience stores, fuel centers and other customer-service venues. Radiant has offices in North America, Europe, Asia and Australia.

NCR Corporation Careers

Join the dynamic team at NCR Corporation, a global leader in consumer transaction technologies, and be part of a company that's redefining the future of connected experiences. At NCR, we believe in fostering a culture of innovation and leadership, making it an ideal place for professionals looking to advance their careers in a vibrant, technology-driven environment.

Work You’ll Do

At NCR Corporation, you’ll have the opportunity to work alongside some of the brightest minds in the industry. Our team is committed to driving growth through innovation and strong leadership, ensuring that we stay ahead of the curve in a rapidly evolving digital landscape. Whether it’s improving consumer interactions, optimizing business processes, or developing the latest in financial technology, your work at NCR will have a lasting impact.

Explore Job Opportunities and Internships

NCR offers a variety of job opportunities and internships across multiple fields and disciplines. From software engineering to sales, customer support to marketing, our positions cater to a wide range of skills and professional interests. Our hiring process is designed to identify and nurture talent, focusing on diversity and the unique contributions of each team member.

Professional Growth and Development

Invest in your future with NCR’s unmatched benefits and career development programs. We provide extensive professional development, including leadership training, skills enhancement workshops, and networking opportunities that pave the way for career advancement. Our commitment to professional growth ensures that every employee has the resources and support they need to succeed.

Diversity and Inclusion

At NCR, we celebrate diversity and are committed to creating an inclusive environment where everyone is respected for their unique perspectives and contributions. Our diversity training programs are integral to our culture, helping us build a stronger, more cohesive team.

Benefits of Joining NCR

Being part of NCR means more than just having a job; it means being part of a community that values your well-being and professional growth. Our comprehensive benefits package includes health, dental, and vision insurance, employee wellness programs, and flexible working arrangements. We understand the importance of work-life balance and strive to provide our employees with an environment that supports both their professional and personal lives.

Stay Connected

Discover the exciting career opportunities at NCR Corporation by exploring our Jobs and Careers page. Tailor your job search with our easy-to-use filters and find the position that best matches your skills and interests. Prepare your resume, sharpen your interview skills, and get ready to embark on a rewarding journey with a company that values innovation and leadership.

Join Our Team

Ready to take the next step in your career? Search open positions at NCR Corporation, update your resume, and apply today. We are continuously looking for passionate, curious, and innovative team players who are ready to make a difference. Join us in shaping the future of connected experiences.

Keep Up to Date

Stay ahead with career tips, insider perspectives, and industry-leading insights you can put to use today—all from the people who work here. Subscribe to our Careers Blog and receive updates that can help propel your career forward at NCR Corporation.

Job Alert Emails

Personalize your subscription to receive job alerts, latest news, and insider tips tailored to your preferences. Discover the exciting and rewarding opportunities that await at NCR Corporation, and see how you can contribute to our culture of growth and innovation.
Learn more about NCR Corporation
Size
38,000 employees
Market Cap
$3.1 billion
Industry
Net Income
-$79 million
Founded
1884
5 Year Trend
+1.8%
Revenue
$6.2 billion
NASDAQ

Similar Jobs

More Jobs at NCR Corporation

More Information Technology Jobs

Find similar Senior Site Reliability Engineer - Unified Observability jobs: