Copy of Senior Data Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+9 more
Job description
Experteer Overview As a Senior Data Engineer at Spokeo, you will build scalable data pipelines and optimize big data workloads using AWS, Spark, and Python. You will collaborate with data science and stakeholders to deliver data products, including entity resolution, and advance our data automation capabilities. You will implement ETL governance, testing, and monitoring to ensure reliable analytics. This remote-first role supports Spokeo’s mission to make data more transparent and actionable. Compensation / Benefits * Build ingestion, processing, and loading data pipelines and automate new components * Collaborate with stakeholders and data science to develop data products including entity resolution * Create unit and stress tests to monitor performance and resolve issues * Develop data analysis tools to extract insights and metrics * Research solutions and maintain technical documentation * Follow data governance, quality, cleansing, and ETL best practices Tasks * 7+ years of data engineering in production environments * Experience with large datasets (>100M records or multi-terabytes) * 5+ years in scalable distributed systems using AWS and EMR * 5+ years Python programming * 5+ years in big data ecosystems; Spark is required (PySpark preferred) * 5+ years SQL, schema design, and dimensional data modeling * 5+ years Airflow or similar dataflow orchestration tools * 2+ years with non-relational databases (e.g., DynamoDB, Elasticsearch) * Bachelor’s degree in Computer Science, Information Systems, Mathematics, or related field Key requirements * bonus program * equity plans * 401(k) * discretionary merit-based salary increases * 100% medical/dental/vision coverage * unlimited employee PTO
Requirements
engineering in production environments * Experience with large datasets (>100M records or multi-terabytes) * 5+ years in scalable distributed systems using AWS and EMR * 5+ years Python programming * 5+ years in big data ecosystems; Spark is required (PySpark preferred) * 5+ years SQL, schema design, and dimensional data modeling * 5+ years Airflow or similar dataflow orchestration tools * 2+ years with non-relational databases (e.g., DynamoDB, Elasticsearch) * Bachelor’s degree in Computer Science, Information Systems, Mathematics, or related field Key requirements * bonus program * equity plans * 401(k) * discretionary merit-based salary increases * 100% medical/dental/vision coverage * unlimited employee PTO
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Data Engineer Salary UK
Making Data Warehouses Fast: A Developer’s Story
Top Big Data Technologies That You Need to Know
Software Engineer Salary London