> Markdown version of [/jobs/ext/586668-sr-data-engineer](https://www.wearedevelopers.com/jobs/ext/586668-sr-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Data Engineer - **Company:** Dentsu Creative - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $94,000.0 - $152,662.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Amazon Web Services, Microsoft Azure, Bash Shell, Unix, Cloud Computing, Code Review, Continuous Integration, Data Deduplication, Information Engineering, Data Governance, Document-Oriented Databases, Java Virtual Machine (JVM), Python (Programming Language), Performance Tuning, Software Engineering, SQL Databases, Data Processing, Large Language Models, Snowflake, Prompt Engineering, Model Validation, Advanced Reports, Git, Software Version Control - **Published:** June 17, 2026 - **Apply:** https://dentsuaegis.wd3.myworkdayjobs.com/DAN_GLOBAL/job/USA---Remote---Maryland/Sr-Data-Engineer_R1124466 ## About the Role * 4+ years of data engineering or software engineering experience, with a focus on data-intensive systems. * Strong Python skills - you write clean, well-structured code and are comfortable building data processing logic from scratch. * Deep Snowflake fluency: data modeling, complex querying, Streams and Tasks, performance tuning, and preferably Snowpark for Python-native workloads. * Strong SQL fundamentals and comfort working with large, messy, real-world datasets - you know how to interrogate data and know when not to trust it. * Some experience or genuine curiosity around identity matching, deduplication, record linkage, or data quality at scale. * Comfort working with PII-class data responsibly, with awareness of data governance and privacy best practices. * Familiarity with version control (Git), Agile delivery, and CI/CD pipelines. * Comfort applying AI tools in day-to-day engineering work - including prompt engineering, LLM-assisted data processing, and AI-augmented pipeline logic. Nice to Have * Hands-on exposure to matching algorithms - deterministic, probabilistic, or ML/AI-based - and experience evaluating or tuning their performance. * Experience building agentic workflows and working with MCP servers * Some Java experience; comfort with JVM-based tooling is a plus. * Familiarity with consumer or household identity signals: name, address, email, phone, and cross-source linkage. * Cloud experience, preferably AWS; Azure or GCP welcome. * Unix/Bash comfort for scripting and day-to-day environment work. ## Description We're looking for an Identity Data Engineer who is passionate about data quality, intellectually curious about how real-world identities get resolved, and ready to get deep into the details. You'll work directly with PII-class data at a low level - examining records, interrogating match logic, and developing a genuine understanding of why our matching engines make the decisions they do. Our matching engines link consumer and household identity signals across diverse data sources, combining deterministic logic with increasingly AI-assisted probabilistic resolution. You'll help enhance these engines - improving match rates, reducing false positives, and extending asset coverage. As our AI-augmented matching capabilities grow, so will this role. There is a real long-term track here for an engineer who wants to go deep on identity. What You'll Do Identity Data Engineering * Design, build, and maintain Snowflake-based pipelines that produce and refresh our core consumer and household identity assets on a regular cadence. * Write complex SQL and Python to transform, deduplicate, and enrich identity data at scale - including direct work with PII fields such as names, addresses, emails, and phone numbers. * Investigate data anomalies and quality issues at a record level, tracing match decisions back to source signals and surfacing root causes. * Build and maintain data models that represent consumer and household identity linkage across multiple input sources. Matching Engine Enhancement * Partner with senior engineers and data scientists to enhance our AI-assisted matching engine - contributing to feature design, scoring logic, model evaluation, and threshold tuning. * Implement and test matching algorithm improvements - both AI-driven and rule-based - and measure their real impact on precision, recall, and overall asset quality. * Build evaluation tooling: ground-truth comparisons, match quality dashboards, and regression detection across engine versions. * Help drive the evolution of our matching pipeline toward more intelligent, AI-augmented identity resolution, actively using AI tools as part of your day-to-day engineering workflow. Collaboration & Delivery * Work cross-functionally with Data Science, Product, and downstream engineering teams to translate identity requirements into reliable, scalable solutions. * Participate in code reviews and architectural discussions; apply engineering best practices across the full delivery lifecycle - design, implement, test, and deploy via CI/CD. * Document data models, pipeline logic, and algorithm decisions clearly for both technical and non-technical audiences. * Support QA processes and on-call responsibilities for production identity asset pipelines. * Build automated validation frameworks and quality tracking pipelines that continuously monitor asset health - including data completeness, match consistency, and anomaly detection - and surface results through clear, actionable reporting. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [WeAreDevelopers LIVE - Node and Package Security](https://www.wearedevelopers.com/videos/2138-wearedevelopers-live-node-and-package-security) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [Hacking AI at the Edge of the Indian Ocean](https://www.wearedevelopers.com/videos/100177-hacking-ai-at-the-edge-of-the-indian-ocean) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)