TikTok

Site Reliability Engineer (Multiple Positions)

TikTok$226K — $316K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Master's degree in Computer Science, Engineering, Information Systems, Mathematics, or related field plus 1 year experience; or Bachelor's degree in the same fields plus 3 years experience.
  • Experience in providing functionality and reliability support for critical site components.
  • Proficient in monitoring system activity and troubleshooting issues.
  • Experience in coordinating data services operations, including SLA management and system deployment.
  • Proven ability to analyze error logs and collaborate on solutions.

Responsibilities

  • Provide site reliability engineering support for large-scale, fault-tolerant systems.
  • Assist in enhancing reliability, scalability, and release cycle of infrastructure services.
  • Measure and monitor availability, latency, and overall service health.
  • Implement sustainable user support, incident response, and blameless postmortems.
  • Develop tools, automation, and monitoring solutions to optimize global infrastructure operation.
  • Establish best design practices for engineers and non-technical team members.
  • Share on-call responsibilities and troubleshoot across various services.

Benefits

  • Day one access to medical, dental, and vision insurance.
  • 401(k) savings plan with company match.
  • Paid parental leave for employees.
  • Short-term and long-term disability coverage available.
  • Life insurance and wellbeing benefits offered.
  • 10 paid holidays and 10 paid sick days per year.
  • 17 days of paid personal time with increasing accruals by tenure.
Full Job Description
Responsibilities

Provide site reliability engineering support to ensure the highest level of availability of large-scale, fault-tolerant systems. Assist the team in improving the reliability, scalability, and release cycle of infrastructure services from inception, design, and development including deployment, user support, and refinement. Measure and monitor availability, latency, and overall service health. Practice sustainable user support, incident response, and blameless postmortems. Build tools, automations, visualizations, and monitors to improve the reliability and scalability of services and to facilitate the operation and optimization of the global infrastructure. Establish solid design and best practices for engineers as well as non-technical team members. Share on-call responsibility and troubleshoot problems across a wide array of services and functional areas.

Qualifications

Qualifications Must have a Master's degree or foreign equivalent degree in Computer Science, Engineering (any), Information Systems, Mathematics, or a related field, and 1 year of related work experience; OR a Bachelor's degree or foreign equivalent degree in Computer Science, Engineering (any), Information Systems, Mathematics, or a related field, and 3 years of related work experience. Of the required experience, must have 1 year of experience in each of the following: Providing functionality and reliability support for critical site components by measuring and monitoring availability, latency, and overall system health, including through performance tuning and troubleshooting; Monitoring system activity and resolving system issues; Coordinating and monitoring data services operations, including SLA management and system deployment; Analyzing error logs to identify issues and working with service owners to resolve issues, document their origins and develop future prevention mechanisms; and Creating and maintaining clear runbook instructions for services to use for alerts, troubleshooting and resolution. Travel Requirement: Domestic and international travel required up to 20%. Type: Full time, 40 hours/week Location: San Jose, CA Salary Range: $226138 - $316800 per year To Apply, click the apply button below. Contact [redacted] if you have difficulty submitting resume through the website.

Job Information

[For Pay Transparency]Compensation Description (Annually)

The base salary range for this position in the selected city is $226138 - $316800 annually.

Compensation may vary outside of this range depending on a number of factors, including a candidate's qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units.

Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure).

The Company reserves the right to modify or change these benefits programs at any time, with or without notice.

About TikTok

TikTok is a social media app that allows users to create and share short videos. The app was launched in 2016 by Chinese tech company ByteDance. TikTok has become one of the most popular social media apps in the world, with over 1 billion active users. The app has been downloaded over 2 billion times worldwide. TikTok has faced controversy over its data privacy practices and its potential ties to the Chinese government. In 2020, the app faced a potential ban in the United States, but a deal was reached with Oracle and Walmart to create a new company called TikTok Global.
Learn more about TikTok
Size
1,750 employees
Industry
Founded
2012

Similar Jobs

More Jobs at TikTok

More Information Technology Jobs

Find similar Site Reliability Engineer (Multiple Positions) jobs: