> Markdown version of [/jobs/ext/525989-lead-data-engineer](https://www.wearedevelopers.com/jobs/ext/525989-lead-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data Engineer - **Company:** Publicis Groupe - **Location:** Chicago, IL, United States - **Experience:** Expert - **Salary:** $98,000.0 - $182,000.0 - **Contract:** Temporary contract - **Skills:** Clean Code Principles, Adobe InDesign, Artificial Intelligence, Airflow, Amazon Web Services, Big Data, Cloud Computing, Code Review, Computer Programming, Computer Engineering, Data Integration, Cursor (Graphical User Interface Elements), Software Debugging, Distributed Systems, Python (Programming Language), NoSQL, Standard Sql, Scala (Programming Language), SQL Databases, Systems Integration, Workflow Management Systems, Data Processing, Cloud Platform System, Concurrency, Apache Spark, Information Technology, Apache Kafka, Video Streaming, Restful APIs, Code Restructuring, Data Pipelines, User Identification, Amazon Elastic Mapreduce (EMR), Databricks - **Published:** June 14, 2026 - **Apply:** https://www.juju.com/job/00000000g7qpaa ## About the Role + S. in Computer Science, Computer Engineering, or a related field. + 5+ years of professional experience on a development team building and maintaining big data pipelines. + Deep, hands-on expertise in Apache Spark and distributed computing concepts. + Strong programming proficiency in Scala and/or Python. + Fluent SQL skills with the ability to ingest complex use cases, refactor code, and tune queries for massive datasets. + Proven experience working within cloud environments (AWS preferred) and managed platforms like Databricks or EMR. + Ability to solve production issues autonomously and own a problem to the end. + Excellent communication skills to work with internal partners, ask the right questions, and translate business requirements into technical solutions. + **Why you might stand out from other talent:** + Experience in the AdTech/MarTech space, specifically programmatic advertising, identity resolution, or audience activation. + Familiarity with orchestration tools (e.g., Airflow). + Experience building and optimizing data integrations with external REST APIs at high concurrency. + Exposure to streaming technologies (Kafka) or NoSQL databases. ## Description The Epsilon Activation Delivery Platform team builds the core framework and connectors powering our audience activation business. We are a small, high-impact team of engineers operating at massive scale-trillions of user events flow through our pipelines every month. By building and maintaining integrations with major advertising players (Meta, Google, Amazon) and Connected TV publishers (LG, Samsung), your work directly connects our data ecosystem to the biggest platforms in the world. Who does this role report to and who will they collaborate with on a day-to-day basis? This role reports to the Engineering Director and collaborates closely with fellow data engineers, product managers, and partner teams. You will participate in design reviews, code reviews, and the end-to-end delivery of critical business features in a highly collaborative environment. Why would a top candidate evaluating multiple opportunities want this job? + **Operate at True Scale:** Work in a high-volume data environment where your pipelines will process and route trillions of records monthly. + **Work with Modern Data Tech:** Get hands-on with an innovative big data stack using Spark, Scala, Python, AWS EMR, and Databricks. + **See Your Impact:** Your code will directly enable our business to interface with the world's largest tech, social media, and CTV platforms. + **Own the Solution:** We are a small team, which means you will have the autonomy to take complex problems from ingestion and specification all the way through to production. What You'll Achieve + **Core Contribution:** Write robust, scalable, and maintainable code using Spark (Scala/Python) and SQL to build out our core framework and data pipelines. + **Pioneer AI-Assisted Engineering:** We are on the cutting edge of AI adoption. You will leverage tools like Cursor and Amazon Q Developer to accelerate coding, debugging, and testing. You'll also actively participate in experimenting with the latest developer workflows for agentic, AI-assisted development, helping define the future of how our team builds software. + **Scale Integrations:** Build, maintain, and optimize the high-concurrency data connectors that feed external publishers and ad networks. + **Optimize & Solve:** Dive deep into complex data processing routines on EMR and Databricks. You will actively fix production issues, tune SQL queries, and solve for performance bottlenecks in a distributed environment. + **Team Mentorship & Quality:** Participate in rigorous code reviews, help enforce engineering standard processes, and mentor mid-level/junior engineers to elevate the team's overall codebase. + **Continuous Improvement:** Build and maintain automated production processing routines that fit seamlessly into our existing scheduled cloud infrastructure. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [ Dev Digest 213: Petrol Prices, Agentic Workflows, AI Skills and CODE100!](https://www.wearedevelopers.com/magazine/718-dev-digest-213-petrol-prices-agentic-workflows-ai-skills-and-code100) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Dev Digest 168: Hacking Postgres, Blocking Meta and Fixing CSS](https://www.wearedevelopers.com/magazine/588-dev-digest-168-hacking-postgres-blocking-meta-and-fixing-css)