Senior Site Reliability Engineer

Multi Media LLC

$169K — $215K *
US-AnywhereRemote in United States
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • STEM degree or relevant experience in SRE, DevOps, or Software Engineering
  • Proficiency in Python or Golang; experience with other languages acceptable
  • Experience running web applications at scale
  • Strong Linux administration and knowledge of Linux internals
  • Good understanding of networking concepts and configurations
  • Experience with database administration and configuration
  • Familiarity with DevOps tools like Terraform, Docker, and Kubernetes

Responsibilities

  • Analyze system performance using telemetry data to identify instabilities
  • Enhance scalability, reliability, and performance via software improvements
  • Develop automation tools for a streamlined DevOps pipeline
  • Design and manage infrastructure in both data center and public cloud
  • Conduct predictive failure analysis and disaster planning
  • Administer databases focusing on uptime and performance
  • Participate in incident response and create postmortem reports
  • Collaborate with engineering teams for cross-functional improvements

Benefits

  • Fully remote work option
  • Comprehensive health, vision, dental, and life insurance paid by the company
  • Long and short-term disability insurance
  • Unlimited paid time off (PTO)
  • Annual company closure for holidays
  • Optional 401k with 5% matching
  • 12 paid holidays per year
  • Paid lunches or stipend for remote work
  • Employee assistance and recognition programs
  • And additional perks and benefits!
Full Job Description
The Role

We are seeking a remote Site Reliability Engineer who will elevate our infrastructure resilience and optimize system performance. As we advance into our next phase of growth, we are searching for someone passionate about driving the enhancement, observability, and automation of our cloud-based infrastructure and fostering innovation and efficiency across our platform.

We're hiring across multiple levels, from intermediate engineers with strong production experience through senior and staff engineers who can lead reliability strategy across systems and teams.

You Will
  • Analyze system performance using APM and distributed telemetry data to identify sources of instability
  • Improve scalability, reliability, and performance through software enhancements and patching
  • Develop tools and automation to streamline the DevOps pipeline
  • Design and manage infrastructure in both data center metal environments and in the public cloud
  • Conduct predictive failure analysis and disaster planning
  • Administer and configure databases and key-value stores with a focus on uptime and performance
  • Analyze complex systems to identify operational surprises and minimize downtime
  • Participate in incident response and produce postmortem reports
  • Collaborate with other engineering teams

Requirements
  • STEM degree and/or relevant experience as a Site Reliability Engineer, Devops Engineer, or SWE
  • Proficiency in Python or Golang. Will also accept experience with other compiled or high level language: C, C#, C++, Java, Rust, etc
  • Experience running Web applications at scale
  • Experience with Web application concepts and frameworks: ORM, MVC architecture, Django, Flask, Laravel, etc
  • Proficiency with Linux administration, Bash shell, and strong knowledge of Linux internals (e.g., filesystems, system calls)
  • Strong networking knowledge (e.g., routing, switching, TCP stack) for both metal and cloud (VPC, Security Groups) environments
  • Experience in database administration and configuration
  • Experience with DevOps tools such as Terraform, Ansible, Docker, Kubernetes, ArgoCD, or Helm
  • Willingness to participate in on-call rotation and respond to monitoring and alerting of core website functions as needed

Benefits
What You'll Get
  • Fair and competitive base salary
  • Fully Remote Optional
  • Health, Vision, Dental, and Life Insurance for you and any dependents, with policy premiums covered by the Company
  • Long & Short term disability insurance
  • Unlimited PTO
  • Annual Year-End Company Closure
  • Optional 401k with 5% matching
  • 12 Paid Holidays
  • Paid Lunches in-office, or if Remote, a $125/week stipend via Sharebite
  • Employee Assistance and Employee Recognition Programs
  • And much more!


The Base Salary range for this position is $169,000 - $215,000 USD. This range reflects base salary only and does not include additional compensation or benefits. The range displayed reflects the minimum and maximum range for a new hire across the US for the posted position. A candidate's specific pay will be determined on a case-by-case basis and may vary based on the candidate's job-related skills, relevant education, training, experience, certifications, and abilities of the candidate, as well as other factors unique to each candidate.

Please note: All offers from Multi Media, LLC are made only after a structured, multi-step recruiting process that includes live interviews, followed by a verbal offer before any written agreement is extended. Official communications, including from interviewers, will only come from email addresses ending in [redacted].com.

Similar Jobs

More Jobs at Multi Media LLC

More Information Technology Jobs

Find similar Senior Site Reliability Engineer jobs: