> Markdown version of [/jobs/ext/1313055-data-engineer](https://www.wearedevelopers.com/jobs/ext/1313055-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** DATA ENGINEERING, LLC - **Location:** San Jose, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $148,750.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Business Analytics Applications, Business Logic, Big Data, Data Architecture, Information Engineering, Data Mapping, Data Systems, Data Warehousing, Cursor (Graphical User Interface Elements), Software Debugging, Distributed Computing Environment, Fault Tolerance, Apache Hadoop, Hadoop Distributed File System, MapReduce, Apache Hive, Python (Programming Language), Performance Tuning, Standard Sql, Data Streaming, Scripting, Google Cloud, Apache Yarn, Multi-Agent Systems, Apache Spark, Build Management, Information Technology, Apache Kafka, Presto, Physical Data Models, Looker Analytics, Data Pipelines, Service Stack - **Published:** July 17, 2026 - **Apply:** https://www.dice.com/job-detail/3484e61d-0819-4ddc-ab9d-54dfd2792e5d ## About the Role * p]:inline" data-streamdown="list-item">8+ years of professional experience as a Data Engineer * p]:inline" data-streamdown="list-item">Extensive SQL skills with the ability to write complex, optimized queries * p]:inline" data-streamdown="list-item">Proficiency in at least one scripting language, with Python preferred * p]:inline" data-streamdown="list-item">Extensive experience with big data technologies such as Hadoop (HDFS, YARN, MapReduce), Hive, Kafka, Spark, Airflow, and Presto/Trino * p]:inline" data-streamdown="list-item">Deep expertise in Apache Spark, including performance tuning, optimization, and building scalable batch and streaming data pipelines * p]:inline" data-streamdown="list-item">Proficiency in data modeling, including designing, implementing, and optimizing conceptual, logical, and physical data models to support scalable and efficient data architectures * p]:inline" data-streamdown="list-item">Experience collaborating with cross-functional teams such as developers, analysts, and operations to deliver results * p]:inline" data-streamdown="list-item">BS in Computer Science or a related field; MS in Computer Science preferred * p]:inline" data-streamdown="list-item">Experience with AWS, Google Cloud Platform, or Looker * p]:inline" data-streamdown="list-item">AI literacy and an AI growth mindset ## Description The mission of Roku's Data Engineering team is to develop a world-class big data platform that enables internal and external customers to leverage data to grow their businesses. Data Engineering works closely with business partners and Engineering teams to collect metrics on existing and new initiatives that are critical to business success. As a Data Engineer working on Device metrics, you will design data models & develop scalable data pipelines to capture different business metrics across different Roku Devices., As a Senior Data Engineer on the Viewer Product Data Engineering team, you will play a pivotal role in designing data models and building scalable pipelines to capture business metrics across Roku devices, Roku-powered TVs, web, and mobile clients. Your work will enable data-driven decisions that shape the content discovery and viewing experience for millions of users worldwide. You will build and maintain highly scalable, fault-tolerant distributed data processing systems that handle tens of terabytes of data ingested daily, as well as a petabyte-scale data warehouse. By delivering trusted, high-quality data products, you will help teams understand which product features resonate most with users, measure their impact, and uncover opportunities to continuously improve the Roku experience. You will also participate in architecture discussions, influence the product roadmap, and take ownership of new projects from inception to delivery. This is an exciting opportunity for a data engineering professional who thrives in a fast-paced environment and is passionate about solving complex data challenges at scale., At Roku, we don't just use AI, we work with it. AI agents and smart tools help power drafts, analysis, and repetitive workflows, while our people bring direction, judgment, and accountability. We're looking for curious, adaptable builders who can show how they've used AI or automation to move faster, raise the bar, and scale their impact. We value your AI skills if you have built fluency across the agentic engineering toolchain - coding harnesses like Claude Code or Cursor, MCP servers, custom skills, or agent frameworks. And you can describe projects where you shipped real work with these tools. You know how to drive an agent, verify its output, and ramp on an unfamiliar codebase with an agent helping you. What you'll be doing * p]:inline" data-streamdown="list-item">Build highly scalable, available, and fault-tolerant distributed data processing systems (batch and streaming) that process tens of terabytes of data daily and manage a petabyte-scale data warehouse * p]:inline" data-streamdown="list-item">Design and build quality data solutions and refine diverse datasets into simplified data models that encourage self-service analytics * p]:inline" data-streamdown="list-item">Develop data pipelines that optimize for data quality and are resilient to poor-quality data sources * p]:inline" data-streamdown="list-item">Own data mapping, business logic, transformations, and data quality standards * p]:inline" data-streamdown="list-item">Perform low-level systems debugging, performance measurement, and optimization on large production clusters * p]:inline" data-streamdown="list-item">Participate in architecture discussions, influence the product roadmap, and take ownership and responsibility for new projects * p]:inline" data-streamdown="list-item">Maintain and support existing platforms and drive evolution to newer technology stacks and architectures, Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Why and when should we consider Stream Processing frameworks in our solutions](https://www.wearedevelopers.com/videos/1085-why-and-when-should-we-consider-stream-processing-frameworks-in-our-solutions) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)