Wonder

Senior Site Reliability Engineer

Wonder$158K — $172K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Proven technical leadership and mentorship abilities.
  • Hands-on experience with Infrastructure as Code (IaC), specifically Terraform and/or Pulumi.
  • Strong familiarity with CI/CD systems like Jenkins and Spinnaker.
  • Deep proficiency in Linux operating systems and performance troubleshooting.
  • Advanced scripting skills using Python and Bash for automation.
  • Robust experience with AWS and operational knowledge of Google Cloud Platform (GCP).
  • Expertise in Docker containerization and configuration management tools.
  • Solid understanding of enterprise logging with tools like Vector, Datadog, and Splunk.
  • Strong security insights related to regulated environments.

Responsibilities

  • Design and secure core infrastructure and deployment platforms.
  • Architect and modernize infrastructure as code across cloud environments.
  • Own and optimize multi-cloud CI/CD deployment pipelines.
  • Harden card payment infrastructure for maximum data security.
  • Automate and manage server image lifecycles using Docker and AMIs.
  • Scale enterprise logging platforms for observability and downtime minimization.
  • Participate in incident response and operational excellence initiatives.

Benefits

  • Competitive salary package including equity and 401K.
  • Multiple medical, dental, and vision plans available.
  • Various additional benefits and perks not explicitly listed.
Full Job Description
About The Opportunity

As a Senior Engineer on the Runtime Automation team, you will design, automate, and secure our core infrastructure, deployment platforms, and CI/CD ecosystems. Our team is responsible for managing our centralized Enterprise Logging Platform (Vector, Datadog, Splunk), engineering robust multi-cloud solutions across AWS and GCP, and driving automated deployment infrastructure via Jenkins and Spinnaker. In this role, you will focus deeply on containerization (Docker), Infrastructure as Code (IaC via Terraform and/or Pulumi), server image management (AMI), and low-level Linux operating system engineering, while ensuring top-tier observability and strict security across our mission-critical card payment environment. This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology ecosystem.

The Impact You Will Make
  • Guide Technical Direction and Innovation: help define architectural strategies, introducing modern platform engineering practices, and driving complex technical decisions across infrastructure and CI/CD platforms.
  • Coach and Mentor Engineers: Foster a culture of technical excellence by mentoring and coaching mid-level and junior engineers, conducting rigorous code and architecture reviews, and promoting continuous knowledge sharing.
  • Scale Infrastructure as Code (IaC): Architect, maintain, and modernize declarative infrastructure across cloud environments using Terraform and/or Pulumi to drive reliability and self-service capabilities.
  • Orchestrate Global CI/CD Infrastructure: Own, scale, and optimize high-throughput multi-cloud deployment pipelines using Jenkins and Spinnaker, ensuring fast, secure, and friction-free software delivery workflows.
  • Secure and Standardize Payment Environments: Architect and harden our card payment infrastructure, ensuring the highest standards of data security, compliance, and proactive vulnerability mitigation.
  • Optimize Infrastructure and Container Lifecycles: Own and scale the automated pipeline for AWS AMIs and Docker containers, using configuration management tools to ensure consistent, secure, and reliable image deployment workflows.
  • Optimize Systems: Troubleshoot and optimize mission-critical infrastructure at scale, leveraging deep Linux internals knowledge and kernel-level tuning to maximize efficiency and stability.
  • Advance Enterprise Logging and Observability: Scale and maintain our data ingestion workflows and centralized logging platform using Vector, Datadog, and Splunk, building robust dashboards and advanced alerting mechanisms to minimize downtime.
  • Foster Operational Excellence: Participate in the team's on-call rotation as an incident responder, leading postmortem discussions and writing automation scripts to eliminate repetitive operational toil.

What You Bring to the Table
  • Technical Leadership and Mentorship: Proven experience or strong aptitude for helping guide technical direction, making sound architectural choices, and successfully coaching/mentoring other engineers.
  • Infrastructure as Code (IaC) Expertise: Hands-on familiarity and engineering experience with modern IaC frameworks, specifically Terraform and/or Pulumi, to manage complex cloud resources.
  • CI/CD and Deployment Expertise: Robust, hands-on familiarity with configuring, scaling, and troubleshooting enterprise CI/CD systems, specifically Jenkins pipelines and Spinnaker deployments.
  • Linux Internals and Systems Engineering: Deep, expert-level proficiency in Linux operating systems with proven experience in low-level system performance troubleshooting.
  • Automation and Advanced Scripting: High proficiency in scripting and software automation utilizing Python and Bash to build internal tools, optimize processes, and eliminate technical overhead.
  • Core Cloud Expertise: Strong engineering experience with AWS services (specifically EC2, Load Balancers, and AMI building) along with robust operational familiarity with Google Cloud Platform (GCP).
  • Containerization and Image Automation: Advanced, hands-on experience with Docker containerization and automated configuration management tools (such as Ansible).
  • Observability and Logging Expertise: Solid command over administering, configuring, and scaling large-scale enterprise logging and monitoring solutions, specifically Vector, Datadog, and Splunk.
  • Security-First Mindset: Strong familiarity with building and maintaining secure infrastructure, particularly within highly regulated or card payment environments.
  • Collaboration and Communication: Excellent problem-solving skills, rigorous attention to detail, and a proven ability to lead cross-functional incident responses and knowledge sharing.
  • Nice-to-Have / Plus: Prior exposure to or experience with Kubernetes-native orchestration and architectures (e.g., EKS, GKE).


Our hybrid model requires 3 days a week in the office. That said, many team members choose to come in more often to take advantage of in-person collaboration and connection. You're welcome-and encouraged-to be in the office up to 5 days a week if it works for you.

#LI-Hybrid

Illinois: $158,500 - $172,000 per year.

New York: $176,000 - $191,000 per year.

Wonder uses geographic-specific salary structures, which means the salary offered may vary depending on where the job is located. The final salary offer will take into account various factors, such as the candidate's skills, education, training, credentials, and experience.

Benefits

We offer a competitive salary package including equity and 401K. Additionally, we provide multiple medical, dental, and vision plans to meet all of our employees' needs as well as many benefits and perks that are not listed.

About Wonder

World of Wonder Productions is an American production company founded in 1991 by filmmakers Randy Barbato and Portsmouth-born Fenton Bailey. Based in Los Angeles, California, the company specializes in documentary television and film productions, with credits including the Million Dollar Listing docuseries, RuPaul's Drag Race, and the documentary films Mapplethorpe: Look at the Pictures and The Eyes of Tammy Faye. Together, Bailey and Barbato have produced programming through World of Wonder for HBO, Bravo, HGTV, Showtime, the BBC, Netflix, and VH1. World of Wonder is perhaps best known for its contributions towards LGBTQ programming, for which they won an Outfest Annual Achievement Award in 2011. Their most well known LGBTQ production is RuPaul's Drag Race, having managed the career of drag queen and titular host RuPaul for many years before this, eventually producing the franchise alongside the majority of its live shows, podcasts, television specials, and conventions.
Learn more about Wonder
Industry
Founded
2015

Similar Jobs

More Jobs at Wonder

More Information Technology Jobs

Find similar Senior Site Reliability Engineer jobs: