> Markdown version of [/jobs/ext/2521288-copy-of-senior-data-engineer](https://www.wearedevelopers.com/jobs/ext/2521288-copy-of-senior-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Copy of Senior Data Engineer - **Company:** Spokeo, Inc. - **Location:** United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Web Services, Data Analysis, Big Data, Information Systems, Information Engineering, Data Governance, Extract Transform Load (ETL), Data Warehousing, Distributed Systems, Amazon DynamoDB, Elasticsearch, Data Flow Control, Python (Programming Language), SQL Databases, Apache Spark, Pyspark, Information Technology, Performance Monitor, Non-relational Database, Data Pipelines - **Published:** August 6, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/copy-of-senior-data-engineer-usa-58832062 ## About the Role engineering in production environments * Experience with large datasets (>100M records or multi-terabytes) * 5+ years in scalable distributed systems using AWS and EMR * 5+ years Python programming * 5+ years in big data ecosystems; Spark is required (PySpark preferred) * 5+ years SQL, schema design, and dimensional data modeling * 5+ years Airflow or similar dataflow orchestration tools * 2+ years with non-relational databases (e.g., DynamoDB, Elasticsearch) * Bachelor's degree in Computer Science, Information Systems, Mathematics, or related field Key requirements * bonus program * equity plans * 401(k) * discretionary merit-based salary increases * 100% medical/dental/vision coverage * unlimited employee PTO ## Description Experteer Overview As a Senior Data Engineer at Spokeo, you will build scalable data pipelines and optimize big data workloads using AWS, Spark, and Python. You will collaborate with data science and stakeholders to deliver data products, including entity resolution, and advance our data automation capabilities. You will implement ETL governance, testing, and monitoring to ensure reliable analytics. This remote-first role supports Spokeo's mission to make data more transparent and actionable. Compensation / Benefits * Build ingestion, processing, and loading data pipelines and automate new components * Collaborate with stakeholders and data science to develop data products including entity resolution * Create unit and stress tests to monitor performance and resolve issues * Develop data analysis tools to extract insights and metrics * Research solutions and maintain technical documentation * Follow data governance, quality, cleansing, and ETL best practices Tasks * 7+ years of data engineering in production environments * Experience with large datasets (>100M records or multi-terabytes) * 5+ years in scalable distributed systems using AWS and EMR * 5+ years Python programming * 5+ years in big data ecosystems; Spark is required (PySpark preferred) * 5+ years SQL, schema design, and dimensional data modeling * 5+ years Airflow or similar dataflow orchestration tools * 2+ years with non-relational databases (e.g., DynamoDB, Elasticsearch) * Bachelor's degree in Computer Science, Information Systems, Mathematics, or related field Key requirements * bonus program * equity plans * 401(k) * discretionary merit-based salary increases * 100% medical/dental/vision coverage * unlimited employee PTO ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)