Cracker Barrel Old Country Store, Inc

Site Reliability Engineer

US-AnywhereRemote in Tennessee, US
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in a relevant field or equivalent experience.
  • 3-5+ years of experience in site reliability engineering or related tech role.
  • Experience supporting high-availability web and cloud-hosted applications.
  • Proficient in monitoring, incident management, and CI/CD pipelines.
  • Familiarity with cloud platforms, infrastructure automation, and microservices.

Responsibilities

  • Design and implement reliability practices for digital platforms.
  • Monitor production systems for proactive issue identification.
  • Provide Tier-2 and Tier-3 support for web and mobile applications.
  • Lead incident response and post-incident reviews.
  • Develop automation tools to enhance operational efficiency.
  • Collaborate with development teams to improve deployment processes.
  • Define and track service health indicators and reliability metrics.

Benefits

  • Competitive annual salary with bonus opportunities.
  • Comprehensive medical, dental, and vision benefits starting on day one.
  • Tuition reimbursement and professional development opportunities.
  • Inclusive culture with onboarding and recognition programs.
  • 401k plan with company matching contributions in 90 days.
Full Job Description
What You'll Do - You'll Make the Moment

  • Design, implement, and maintain reliability practices that improve availability, performance, scalability, resiliency, and operational maturity across digital platforms and supporting services.
  • Monitor production systems using observability tools, dashboards, logs, metrics, traces, alerts, and synthetic monitoring to identify issues before they impact customers or associates.
  • Provide Tier-2 and Tier-3 production support for customer-facing and internal digital applications, including web, mobile, commerce, CMS, APIs, integrations, and cloud-hosted services.
  • Lead and participate in incident response, root-cause analysis, problem management, post-incident reviews, and follow-up actions that reduce recurrence and improve service reliability.
  • Develop automation, scripts, runbooks, self-healing processes, and operational tools that reduce manual effort, accelerate recovery, and improve consistency across environments.
  • Partner with development teams to improve CI/CD pipelines, deployment readiness, release validation, rollback procedures, feature monitoring, and environment stability.
  • Collaborate with infrastructure, cloud, security, architecture, QA, and vendor teams to ensure systems meet company standards for security, privacy, compliance, resiliency, and operational support.
  • Define and track service health indicators such as availability, latency, error rates, capacity, incident trends, deployment quality, and other reliability metrics.
  • Create and maintain technical documentation, operational support guides, escalation paths, production readiness checklists, and disaster recovery procedures.
  • Understand and comply with all company privacy, security, accessibility, change management, and technology standards.
  • Bachelor's degree in Computer Science, Computer Information Systems, Software Engineering, Information Technology, or a related discipline is preferred; equivalent experience or training may be considered.
  • 3-5+ years of experience in site reliability engineering, DevOps, cloud operations, production support, systems engineering, software engineering, or a related technology operations role.
  • Experience supporting high-availability web, mobile, commerce, API, integration, or cloud-hosted application environments.
  • Hands-on experience with monitoring, logging, alerting, incident management, root-cause analysis, CI/CD pipelines, Git-based workflows, and release support.
  • Experience with cloud platforms, containers, infrastructure automation, scripting, APIs, microservices, content management systems, or restaurant/retail technology environments preferred.

Focus on You

We're all about making sure you're taken care of too. Here's what's in it for you:
  • Good Work Deserves Good Pay: Competitive Annual Salary | Annual Bonus Opportunities
  • Support That Goes Beyond the Clock: Medical, Rx, Dental and Vision Benefits on Day 1| Life Insurance and Disability Coverage | Paid Vacation/Employee Assistance Program
  • Grow and Thrive Your Way: Business Resource Groups | Tuition Reimbursement | Professional Development
  • Culture of Belonging:Support that starts on day one | Onboarding, training, and development to help you thrive | Recognition programs and employee events that bring us together
  • Invest in Your Future: 401k Plan with Company Matching Contributions at 90 days | Employee Stock Purchase Program
  • More Perks, Just Because: 35% Discount on Cracker Barrel Food and Retail items | Exclusive Biscuit Perks like discounts on home, travel, cell phones, and more!

About Cracker Barrel Old Country Store, Inc

Cracker Barrel Old Country Store, Inc. is a chain of restaurants with a Southern country theme. The company operates 664 Cracker Barrel locations in 45 states. The restaurants serve breakfast, lunch, and dinner, and offer a retail store that sells various items, including rocking chairs, holiday and seasonal gifts, apparel, toys, music CDs, cookware, and various other gift items. Cracker Barrel's menu features home-style country food, including biscuits, chicken and dumplings, fried chicken, meatloaf, and pancakes. Cracker Barrel generated $2.3 billion in revenue in 2020.
Learn more about Cracker Barrel Old Country Store, Inc
Size
70,000 employees
Market Cap
$2.1 billion
Industry
Net Income
$47.8 million
5 Year Trend
+2.2%
Revenue
$2.2 billion
NASDAQ

Similar Jobs

More Jobs at Cracker Barrel Old Country Store, Inc

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: