Scientific Games Corporation

Sr. Site Reliability Engineer (SRE)

Scientific Games Corporation$120K — $145K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor’s degree in computer science or related field, or equivalent work experience.
  • 6+ years of experience as an SRE, DevOps Engineer, or similar role.
  • Strong experience with AWS (EKS, EC2, S3, Route53, IAM).
  • 6+ years managing production Kubernetes workloads.
  • Hands-on experience with New Relic, Graylog, or similar monitoring tools.
  • Experience with HashiCorp Vault or equivalent for secrets management.
  • Proficiency with GitHub Actions, GitLab CI/CD, Helm, and ArgoCD.
  • Hands-on experience with Terraform and proficiency in Python, Bash, or equivalent scripting languages.

Responsibilities

  • Enhance observability using New Relic, Graylog, or similar tools.
  • Establish actionable alerting and dashboards for service health metrics.
  • Implement and maintain reliable systems for high availability and scalability.
  • Define and monitor SLIs, SLOs, and SLAs with cross-functional teams.
  • Automate operational processes to reduce manual interventions.
  • Manage Kubernetes workloads on AWS EKS for secure deployments.
  • Participate in on-call rotation, troubleshoot incidents, and lead post-incident reviews.

Benefits

  • Work within a highly regulated gaming and lottery environment.
  • Opportunity to enhance technical skills with cutting-edge tools.
  • Participate in cross-functional collaborations with development and DevOps teams.
  • Gain exposure to incident management and post-incident analysis.
  • Work with modern cloud technologies and automation frameworks.
Full Job Description
Position Summary

We are looking for a skilled Site Reliability Engineer (SRE) to enhance the stability, performance, and reliability of our production systems. The SRE will work closely with development, DevOps, and security teams, ensuring production readiness, managing on-call responsibilities, and improving observability across applications and infrastructure.

  • Monitoring & Observability
    • Maintain and enhance observability using New Relic, Graylog, OR other monitoring tools.
    • Establish actionable alerting and dashboards for service health and performance metrics.
  • Reliability Engineering
    • Implement and maintain reliable systems, focusing on capacity planning, performance optimization, and fault tolerance to ensure high availability and scalability.
    • Collaborate with teams to define and implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs), and monitor their performance.
  • Automation & Infrastructure Operations
    • Automate operational processes, reducing manual interventions.
    • Manage Kubernetes workloads on AWS EKS, ensuring secure and stable deployments.
    • Work with HashiCorp Vault for secrets management and security compliance.
  • Incident & Problem Management
    • Participate in on-call rotation to handle production incidents and ensure rapid resolution.
    • Troubleshoot production issues, identify root causes, and implement permanent fixes.
    • Lead post-incident reviews, create action items, and follow through on remediation.
  • Collaboration
    • Work closely with DevOps to improve CI/CD pipelines for production readiness.
    • Partner with development teams to embed resilience and observability into applications.
  • Documentation & Knowledge Sharing
    • Document operational runbooks, escalation procedures, and production playbooks.

Qualifications

Required Skills

  • Bachelor’s degree in computer science or related field, or equivalent work experience.
  • Experience: 6+ years as an SRE, DevOps Engineer, or similar role
  • Cloud: Strong experience with AWS (EKS, EC2, S3, Route53, IAM)
  • Kubernetes: 6+ years managing production Kubernetes workloads
  • Monitoring & Observability: Hands-on with New Relic, Graylog, or similar
  • Secrets Management: Experience with HashiCorp Vault or equivalent
  • Automation & CI/CD: Proficiency with GitHub Actions, GitLab CI/CD, Helm and ArgoCD
  • IaC : Hands-on experience with Terraform
  • Scripting: Proficiency in Python, Bash, or equivalent scripting languages
  • Incident Management: Strong debugging, troubleshooting, and root cause analysis skills
  • On-Call Readiness: Willingness to participate in 24x7 on-call rotation

Desired Skills

  • AWS certification
  • Familiarity with .NET application stack
  • Multi-cloud exposure
  • Experience managing Kubernetes clusters with Rancher in on-prem environments
  • Familiarity with Packer for building Golden AMIs

Physical Requirements

The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. While performing the duties of this job, the employee is regularly required to sit, stand, walk, bend, use hands, operate a computer, and have specific vision abilities to include close and distance vision, and ability to adjust focus working with computer and business equipment.


Work Conditions

Scientific Games Corporation and its affiliates (collectively, “SG”) are engaged in highly regulated gaming and lottery businesses.   As a result, certain SG employees may, among other things, be required to obtain a gaming or other license(s), undergo background investigations or security checks, or meet certain standards dictated by law, regulation or contracts.   In order to ensure SG complies with its regulatory and contractual commitments, as a condition to hiring and continuing to employ its employees, SG requires all of its employees to meet those requirements that are necessary to fulfill their individual roles.  As a prerequisite to employment with SG (to the extent permitted by law), you shall be asked to consent to SG conducting a due diligence/background investigation on you.

This job description should not be interpreted as all-inclusive; it is intended to identify major responsibilities and requirements of the job. The employee in this position may be requested to perform other job-related tasks and responsibilities than those stated above. 

About Scientific Games Corporation

Light & Wonder, Inc., formerly Scientific Games Corporation, is an American corporation that provides gambling products and services. The company is headquartered in Las Vegas, Nevada, with lottery headquarters and production plant in Alpharetta, Georgia. Light & Wonder's gaming division provides products such as slot machines, table games, shuffling machines, and casino management systems. Its brands include Bally, WMS, and Shuffle Master.
Learn more about Scientific Games Corporation
Size
9,500 employees
Market Cap
$5.6 billion
Industry
Net Income
-$569 million
Founded
1973
5 Year Trend
-5.7%
Revenue
$2.7 billion
NASDAQ

Similar Jobs

More Jobs at Scientific Games Corporation

More Information Technology Jobs

Find similar Sr. Site Reliability Engineer (SRE) jobs: