The Test Automation Platform team (TAP) provides the core infrastructure and capabilities to enable automated testing of the Netflix product at scale. Our Device & Test Automation platform is used to enable other teams to qualify and validate the Netflix TV, mobile, and web client applications, partner device implementations, mobile games, and more. We view ourselves as a force multiplier for Netflix engineering, providing composable capabilities and pluggable abstractions that allow teams to manage, orchestrate, and analyze their automated tests and devices. Our platform executes and ingests results for over 3 million test executions daily.
You will join the Infrastructure & Operations pod within TAP. This pod develops and operates the foundational services and infrastructure that underpin the Test Automation platform, covering service deployment and delivery, core datastores, observability and alerting, CI/CD and developer tooling, and the reliability and efficiency of the platform as it scales.
What you will do:
- Design, build, and operate backend services and infrastructure that power the Test Automation platform, with a focus on reliability, scalability, and cost efficiency.
- Own and evolve core infrastructure components such as service deployment and delivery, platform datastores (Aurora, MongoDB, DocumentDB).
- Standardize and modernize service infrastructure by moving services onto paved paths for observability, provisioning, capacity management, security.
- Analyze and optimize critical systems like MongoDB for capacity, performance, and cost, including sharding, version upgrades, and data lifecycle strategies (TTL, archival, hot/cold storage).
- Improve operational excellence by tuning metrics and alert sources, enhancing dashboards and alerts, and building runbooks.
- Drive resilience and reliability initiatives such as load testing, failure injection testing, disaster recovery and high-availability strategies, and post-incident improvements grounded in cost/benefit tradeoffs.
- Collaborate with partner teams on paved paths, storage and compute offerings, and ETL/data pipelines.
This role is right for you if:
- You think in terms of platforms and paved paths and building opinionated, reusable infrastructure components that other engineers can rely on.
- You have a strong infrastructure and backend engineering background and enjoy working across services, storage, compute, and operations for large-scale platforms.
- You have experience working with cloud provider technologies (AWS, GCP, Azure)
- You have experience with distributed systems fundamentals such as latency, throughput, backpressure, retries, idempotency, and consistency and availability tradeoffs.
- You have significant experience in at least one programming language used in backend development (Java, Golang, C++, Typescript).
- You understand networking fundamentals and can debug issues involving TCP/IP, HTTP, TCP, DNS, etc.
- You take initiative and drive projects with dedication.
- You excel in collaborative settings and use your strong communication skills to influence outcomes.
Bonus Skills:
- Experience with MongoDB or DocumentDB sharding, indexing, performance tuning, data lifecycle management.
- Experience with infrastructure for large-scale test or CI systems, including scheduling, queuing, parallel execution, and resource-aware scaling.
- Experience building or operating data pipelines and ETL from operational data stores into analytics systems such as Iceberg, data lakes, or data warehouses.
- Familiarity with resilience engineering practices, including failure injection, DR and multi-region strategies, incident reviews, and Linux systems debugging from the command line.
Generally, our compensation structure consists solely of an annual salary; we do not have bonuses. You choose each year how much of your compensation you want in salary versus stock options. To determine your personal top of market compensation, we rely on market indicators and consider your specific job family, background, skills, and experience to determine your compensation in the market range. The range for this role is $388,000.00 - $558,000.00.
Netflix provides comprehensive benefits including Health Plans, Mental Health support, a 401(k) Retirement Plan with employer match, Stock Option Program, Disability Programs, Health Savings and Flexible Spending Accounts, Family-forming benefits, and Life and Serious Injury Benefits. We also offer paid leave of absence programs. Full-time hourly employees accrue 35 days annually for paid time off to be used for vacation, holidays, and sick paid time off. Full-time salaried employees are immediately entitled to flexible time off. See more details about our Benefits here.
Netflix is a unique culture and environment. Learn more here.