OpenTable

Site Reliability Engineer II (AI Platform)

OpenTable • $110K — $130K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of hands-on Linux experience (e.g., Ubuntu, CentOS) with expertise in kernel tuning and performance optimization.
  • 3+ years of experience with configuration management systems like Puppet, Chef, Ansible, or SaltStack.
  • Proven experience building and troubleshooting bare-metal Kubernetes clusters, including control plane management.
  • Proficiency in automation and scripting languages such as Go, Python, Ruby, Perl, or Bash.
  • Experience in leading root-cause analysis for live service disruptions and operating messaging systems in production.

Responsibilities

  • Maintain high availability and performance of Linux OS and Kubernetes across bare-metal infrastructure.
  • Architect and manage scalable container and configuration management tools globally.
  • Conduct root-cause analysis for infrastructure disruptions and performance bottlenecks.
  • Collaborate with distributed engineering teams on high-impact platform projects.
  • Support critical production systems through an on-call rotation.
  • Develop self-service tools and automation to reduce operational overhead.

Benefits

  • Work remotely for up to 20 days per year.
  • Company-paid therapy and subscription to Headspace for mental health.
  • Annual company-wide week off for team recharge.
  • Generous paid vacation and time off for your birthday.
  • Paid parental leave and volunteer time.
  • Career growth opportunities with development dollars and leadership support.
  • 20 days of paid time off and comprehensive health and dental insurance.
  • Life and disability insurance.
Full Job Description
This hybrid role requires working in the office two days per week.

About the job

As a Site Reliability Engineer II on the Serving Platforms team within Infrastructure Engineering, you will design, automate, and manage the core container stack and infrastructure powering our global business applications. Operating in a high-scale, self-hosted environment, you will serve as a subject matter expert for Kubernetes, Linux systems, and cloud-native automation, directly driving the reliability, security, and efficiency of our platform. In this role, you will collaborate with cross-functional engineering teams worldwide, lead greenfield infrastructure projects, resolve complex incidents, and build self-service capabilities that empower application developers across the organization.

Responsibilities
  • Maintain, tune, and ensure high availability for the low-level Linux operating system and Kubernetes control plane across our self-hosted bare-metal infrastructure.
  • Architect, build, and maintain scalable container management, configuration management, and automation tools across global environments.
  • Investigate, resolve, and conduct root-cause analysis for complex infrastructure disruptions and performance bottlenecks at the system call level.
  • Participate in high-impact platform engineering projects and collaborate with globally distributed engineering teams to drive infrastructure standardization.
  • Participate in the team's on-call rotation to support critical production systems and ensure operational resilience.
  • Develop and maintain self-service tools, automation pipelines, and robust infrastructure monitoring to eliminate manual operational overhead.

Minimum Qualifications
  • 5+ years of hands-on Linux experience (e.g., Ubuntu, CentOS) with expertise in kernel tuning (sysctl), process management (cgroups/namespaces), system calls, and performance optimization.
  • 3+ years of experience using configuration management systems such as Puppet, Chef, Ansible, or SaltStack in production environments.
  • Proven experience building, operating, and troubleshooting bare-metal Kubernetes clusters from the ground up, including control plane, etcd, and CNI plugin management.
  • Proficiency with continuous system automation and scripting in languages such as Go, Python, Ruby, Perl, or Bash.
  • Demonstrated experience responding to live service disruptions, leading root-cause analysis, and operating messaging systems (e.g., Kafka or RabbitMQ) in production.

Preferred Qualifications
  • Experience operating, scaling, and monitoring AI/ML or LLM-powered services and workloads in high-concurrency production environments.
  • Hands-on expertise with public cloud providers (AWS, GCE, or Azure) and containerized CI/CD pipelines (e.g., GitHub, Jenkins, CircleCI, Docker).
  • Experience with distributed key-value stores (e.g., Consul, etcd, Zookeeper, Redis) and enterprise observability/alerting tools (e.g., Prometheus, Sensu).
  • Familiarity with server virtualization infrastructure (e.g., Proxmox, VMware, Xen, OpenStack) and low-level networking concepts (IPtables/NFTables, routing, load balancing).
  • Experience developing and maintaining OS-level software packaging (RPM/DEB) and participating in globally distributed software engineering teams.

This posting is for an existing vacancy.
Benefits and Perks
  • Work from (almost) anywhere for up to 20 days per year
  • Focus on mental health and well-being:
    • Company-paid therapy sessions through SpringHealth
    • Company-paid subscription to Headspace
    • Annual company-wide week off a year - the whole team fully recharges (and returns without a pile-up of work!)
  • Paid parental leave
  • Generous paid vacation + time off for your birthday
  • Paid volunteer time
  • Focus on your career growth:
    • Development Dollars
    • Leadership development
    • Access to thousands of on-demand e-learnings
  • Travel Discounts
  • Employee Resource Groups
  • 20 days of paid time off
  • Private health and dental insurance
  • Life and Disability insurance

The best connections happen face-to-face, whether you're sitting down to dinner or having coffee with a coworker. That's why OpenTable has adopted a hybrid workplace model. This role aligns with that approach, with an expectation of coming into the office two days a week-giving employees the best of both worlds: in-person collaboration and flexibility.

The best connections happen face-to-face, whether you're sitting down to dinner or having coffee with a coworker. That's why OpenTable has adopted a hybrid workplace model. This role aligns with that approach, with an expectation of coming into the office two days a week-giving employees the best of both worlds: in-person collaboration and flexibility.

The expected range of compensation for this position based in Toronto, Canada, including commission and/or bonuses, is $110,000-$130,000 CAD. There are a variety of factors that go into determining a compensation range, including but not limited to external market benchmark data, geographic location, and years of experience sought/required.

We offer a competitive base salary and benefits including: health benefits; flexible spending account; retirement benefits; life insurance; paid time off (including PTO, paid sick leave, medical leave, bereavement leave, floating holidays and paid holidays); and parental leave benefits. This role is eligible to be considered for an annual bonus and equity grant.

Work Environment & Flexibility

At OpenTable, we pride ourselves on fostering a global and dynamic work environment. As a team member with us, you will benefit from a schedule tailored to accommodate a global workforce operating across multiple time zones. While the majority of your responsibilities may align with conventional business hours, there will be instances where you are expected to manage communications - via calls, Slack messages, or emails - outside of regular working hours to effectively collaborate with international colleagues, respond to restaurant partners, and/or address urgent matters. OpenTable will always abide by and consider local laws and regulations.

Inclusion

We're committed to creating a workplace where everyone feels they belong and can thrive. We know the best ideas come when we bring different voices to the table, so we're building a team as dynamic as the diners and restaurants we serve-and fostering a culture where everyone feels welcome to be themselves.

If you need accommodations during the application or interview process, or on the job, we're here to support you. Please reach out to your recruiter to request any accommodations.

#LI-Hybrid

About OpenTable

OpenTable is an online restaurant-reservation service company. It provides online reservations at about 54,000 restaurants around the world seating some 30 million diners monthly. The company was founded in San Francisco, California in 1998. Reservations are free to end users; the company charges restaurants monthly and per-reservation fees for their use of the system. In addition to online reservations, OpenTable offers restaurant management software for table management, guest management, and reservation management. OpenTable went public on the NASDAQ in 2009, and in 2014 it was acquired by Priceline Group for $2.6 billion.
Learn more about OpenTable
Size
5,000 employees
Industry
Founded
1998

Similar Jobs

More Jobs at OpenTable

More Information Technology Jobs

Find similar Site Reliability Engineer II (AI Platform) jobs: