Software Engineer 5 - Developer Infrastructure & Quality Platform

Netflix, Inc.
Los Gatos, United States
2 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Amazon Web Services Automation of Tests Microsoft Azure C++ (Programming Language) Command-Line Interface Extract Transform Load (ETL) Data Systems Data Warehousing Shard (Database Architecture) Software Debugging Linux
+21 more
Disaster Recovery Distributed Systems Domain Name System (DNS) Hypertext Transfer Protocols (HTTP) Information Lifecycle Management Load Testing MongoDB Networking Basics Operational Data Store Performance Tuning TCP/IP Transmission Control Protocol (TCP) TypeScript Google Cloud Indexer Backend Data Lakes Low Latency Data Pipelines Golang Programming Languages

Job description

  • Design, build, and operate backend services and infrastructure that power the Test Automation platform, with a focus on reliability, scalability, and cost efficiency.
  • Own and evolve core infrastructure components such as service deployment and delivery, platform datastores (Aurora, MongoDB, DocumentDB).
  • Standardize and modernize service infrastructure by moving services onto paved paths for observability, provisioning, capacity management, security.
  • Analyze and optimize critical systems like MongoDB for capacity, performance, and cost, including sharding, version upgrades, and data lifecycle strategies (TTL, archival, hot/cold storage).
  • Improve operational excellence by tuning metrics and alert sources, enhancing dashboards and alerts, and building runbooks.
  • Drive resilience and reliability initiatives such as load testing, failure injection testing, disaster recovery and high-availability strategies, and post-incident improvements grounded in cost/benefit tradeoffs.
  • Collaborate with partner teams on paved paths, storage and compute offerings, and ETL/data pipelines.

Requirements

  • You have a strong infrastructure and backend engineering background and enjoy working across services, storage, compute, and operations for large-scale platforms.
  • You have experience working with cloud provider technologies (AWS, Google Cloud Platform, Azure)
  • You have experience with distributed systems fundamentals such as latency, throughput, backpressure, retries, idempotency, and consistency and availability tradeoffs.
  • You have significant experience in at least one programming language used in backend development (Java, Golang, C++, Typescript).
  • You understand networking fundamentals and can debug issues involving TCP/IP, HTTP, TCP, DNS, etc.
  • You take initiative and drive projects with dedication.
  • You excel in collaborative settings and use your strong communication skills to influence outcomes.

Bonus Skills:

  • Experience with MongoDB or DocumentDB sharding, indexing, performance tuning, data lifecycle management.
  • Experience with infrastructure for large-scale test or CI systems, including scheduling, queuing, parallel execution, and resource-aware scaling.
  • Experience building or operating data pipelines and ETL from operational data stores into analytics systems such as Iceberg, data lakes, or data warehouses.
  • Familiarity with resilience engineering practices, including failure injection, DR and multi-region strategies, incident reviews, and Linux systems debugging from the command line.

Benefits & conditions

  • Netflix’s unique culture is not just a memo but something we practice daily. A description of our culture and way of working may be found here.
  • You will work with an amazing and passionate team of high-performing colleagues invested in your success.
  • Your work will be a force multiplier for hundreds of engineers across Netflix and impacts the Netflix user experience which is used by tens of millions of people every day. When your friends and family ask you what you do for a living, you can point to Netflix running on their TV and say with pride, “I make that possible.”

Generally, our compensation structure consists solely of an annual salary; we do not have bonuses. You choose each year how much of your compensation you want in salary versus stock options. To determine your personal top of market compensation, we rely on market indicators and consider your specific job family, background, skills, and experience to determine your compensation in the market range. The range for this role is $388,000.00 - $558,000.00.

Netflix provides comprehensive benefits including Health Plans, Mental Health support, a 401(k) Retirement Plan with employer match, Stock Option Program, Disability Programs, Health Savings and Flexible Spending Accounts, Family-forming benefits, and Life and Serious Injury Benefits. We also offer paid leave of absence programs. Full-time hourly employees accrue 35 days annually for paid time off to be used for vacation, holidays, and sick paid time off. Full-time salaried employees are immediately entitled to flexible time off. See more details about our Benefits here.

Netflix is a unique culture and environment. Learn more here.

About the company

At Netflix, our mission is to entertain the world. Together, we are writing the next episode - pushing the boundaries of storytelling, global fandom and making the unimaginable a reality. We are a dream team obsessed with the uncomfortable excitement of discovering what happens when you merge creativity, intuition and cutting-edge technology. Come be a part of what’s next.

The Test Automation Platform team (TAP) provides the core infrastructure and capabilities to enable automated testing of the Netflix product at scale. Our Device & Test Automation platform is used to enable other teams to qualify and validate the Netflix TV, mobile, and web client applications, partner device implementations, mobile games, and more. We view ourselves as a force multiplier for Netflix engineering, providing composable capabilities and pluggable abstractions that allow teams to manage, orchestrate, and analyze their automated tests and devices. Our platform executes and ingests results for over 3 million test executions daily.

You will join the Infrastructure & Operations pod within TAP. This pod develops and operates the foundational services and infrastructure that underpin the Test Automation platform, covering service deployment and delivery, core datastores, observability and alerting, CI/CD and developer tooling, and the reliability and efficiency of the platform as it scales.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:02 min

Mapping distributed compute paradigms to modern vehicles

Joachim Werner · LIVE

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

3:36 min

Critical infrastructure and performance skills for modern developers

Andrew Holway · LIVE

3:50 min

Queues in TCP stacks and continuous network connections

Clemens Vasters Clemens Vasters · World Congress 2022

Videos

See all

Related articles

See all