Senior Site Reliability Engineer (Multiple Positions)

ByteDance

$212K — $368K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Master's degree in Computer Science, Engineering, Information Systems, Data Science, Mathematics, or related field and 2 years of relevant experience, OR Bachelor's degree in similar fields and 5 years of post-bachelor's experience.
  • 2 years of experience providing functionality and reliability support for critical site components.
  • Experience working across all phases of the SDLC, including requirements gathering, design, development, and deployment.
  • Proven ability to create and maintain clear runbook instructions for services.
  • 2 years of Linux administration experience, including performance monitoring and troubleshooting.

Responsibilities

  • Provide site reliability engineering support for high availability of large-scale systems.
  • Deliver tools and software to enhance the reliability and scalability of services.
  • Measure and monitor system availability and performance metrics.
  • Practice sustainable incident response and execute postmortems after outages.
  • Establish best practices for design and operations for technical and non-technical staff.
  • Mentor junior engineers and interns to foster professional growth.

Benefits

  • Medical, dental, and vision insurance accessible from day one.
  • 401(k) savings plan with company match.
  • Paid parental leave available.
  • Disability coverage for both short and long term.
  • Life insurance and wellbeing benefits offered.
  • Ten paid holidays and ten paid sick days each year.
  • Seventeen days of paid personal time, increasing with tenure.
Full Job Description
Responsibilities Provide site reliability engineering support to ensure highest level of availability of large-scale, fault tolerant systems. Deliver tools and software to improve the reliability, scalability and operability of services, including designing, developing and deploying automation to sustainably scale with quality. Measure and monitor availability, latency and overall service health. Practice sustainable incident response and postmortems, performing root cause analysis of incidents to influence future product design and response activities. Establish solid design and best practices for engineers as well as non-technical team members. Mentor junior engineers and intern. Qualification Qualifications Must have a Master's degree or foreign equivalent degree in Computer Science, Engineering (any), Information Systems, Data Science, Mathematics, or a related field, and 2 years of related work experience; OR a Bachelor's degree or foreign equivalent degree in Computer Science, Engineering (any), Information Systems, Data Science, Mathematics, or a related field, and 5 years of post-bachelor's, progressive related work experience. Of the required experience, must have 2 years of experience in each of the following: Providing functionality and reliability support for critical site components by measuring and monitoring availability, latency, and overall system health; Working across all phases of the SDLC, including requirements gathering and analysis, design, development, implementation, testing, deployment, and maintenance of back-end and cloud native projects; Creating and maintaining clear runbook instructions for services to use for alerts, troubleshooting and resolution; Coordinating and monitoring data services operations, including SLA management and system deployment; and Performing Linux administration, including monitoring performance, debugging issues, monitoring network behavior, and troubleshooting networked applications using: OS networking protocol stack and the following OS concepts: virtualization and containerization. Travel Requirement: International and domestic travel required up to 10%. Type: Full time, 40 hours/week Location: Bellevue, WA Salary Range: $212202 - $368220 per year To Apply, click the apply button below. Contact [redacted] if you have difficulty submitting resume through the website. Job Information 【For Pay Transparency】Compensation Description (Annually) The base salary range for this position in the selected city is $212202 - $368220 annually. Compensation may vary outside of this range depending on a number of factors, including a candidate's qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units. Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure). The Company reserves the right to modify or change these benefits programs at any time, with or without notice.

Similar Jobs

More Jobs at ByteDance

More Information Technology Jobs

Find similar Senior Site Reliability Engineer (Multiple Positions) jobs: