> Markdown version of [/jobs/ext/566616-data-engineer](https://www.wearedevelopers.com/jobs/ext/566616-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Dow - **Location:** Midland, MI, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Query Performance, Airflow, Data Analysis, Microsoft Azure, Cloud Computing, Code Review, Databases, Continuous Delivery, Continuous Integration, Information Engineering, Data Governance, Data Integration, Data Integrity, Data Transformation, Persistent Data Structure, Data Systems, Data Visualization, Dimensional Modeling, Distributed Computing Environment, Graph Database, Microsoft SQL Server, Neo4j, Performance Tuning, Power BI, Software Engineering, Tableau (Software), Azure Data Factory, Sql Optimization, Apache Spark, Software Security, Git, Pyspark, Star Schema, Cosmos DB, Terraform, Data Pipelines, Databricks - **Published:** June 5, 2026 - **Apply:** https://www.juju.com/job/00000000g5k9zk ## About the Role A successful candidate will possess the experience and technical depth required to independently implement and optimize complex data solutions: + Core Technical Expertise (2-5 Years Demonstrated Experience) + PySpark and Distributed Processing: Proven ability to write highly optimized, production-grade PySpark/Spark code. Experience identifying and resolving performance bottlenecks in a distributed computing environment. + Advanced Data Modeling: Practical experience designing and implementing analytical data models (e.g., dimensional modeling, star/snowflake schemas) and handling Slowly Changing Dimensions (SCDs). + Cloud Orchestration: Expertise in using Azure Data Factory (ADF), Databricks Workflows, or equivalent tools (e.g., Airflow) for complex dependency management, error handling, and end-to-end pipeline orchestration. + Database Versatility: Demonstrated experience with advanced SQL and hands-on experience querying and integrating data from at least one non-relational or Graph database (e.g., CosmosDB, Neo4j). + Engineering Mindset and Professional Growth + Technical Design Contribution: Ability to rapidly synthesize information and contribute clear, well-documented technical specifications and architectural diagrams to the design process. + Feature Ownership: Demonstrated history of taking ownership of complex features and modules within larger projects, driving them to completion, and managing technical dependencies autonomously. + Pragmatism and Initiative: A strong bias for action, coupled with a pragmatic approach to delivering stable, maintainable, and cost-effective solutions. + Communication & Influence: Excellent verbal and written communication skills, with the ability to articulate technical designs to both engineering peers and senior stakeholders, effectively influencing technical decisions. Required Qualifications + A minimum of a bachelor's degreeorrelevant military experience at or above a U.S. E5 rankingorCanadian Petty Officer 2nd Class or Sergeant OR 5 years relevant experience in lieu of a Bachelor's degree. + Minimum of 2 years of professional experience in Data Engineering, Software Engineering, or a closely related field. + Minimum of 2 years ofhands-onexperience with Databricks Platform. + A minimum requirement for this U.S. based position is the ability to work legally in the United States. No visa sponsorship/supportis available for this position, including for any type of U.S. permanent residency (green card) process. Preferred Skills + Experience with cloud cost management principles related to compute (Databricks) and storage (ADLS). + Experience with Infrastructure as Code (e.g., Terraform, ARM templates). + Proficiency with data visualization and dashboarding tools (e.g., Power BI, Tableau). Your Skills + PySpark / Distributed Data Processing:The ability to build, optimize, and troubleshoot high-volume data transformation pipelines using PySpark, including tuning Spark jobs, resolving performance bottlenecks, and ensuring efficient distributed execution. + Advanced Data Modeling (Dimensional / Star Schema Design):Expertise in translating complex business requirements into scalable analytical data models-such as star and snowflake schemas-and implementing SCD logic for downstream analytics. + Cloud Orchestration & CI/CD (Azure Data Factory, Databricks Workflows, Azure DevOps/Git):Skill in designing automated, reliable data pipelines, managing task dependencies, and implementing CI/CD deployment processes across environments. + Data Integration Across Diverse Systems (SQL Server, CosmosDB, Neo4j):Ability to connect to, query, and integrate data from relational and non-relational sources while optimizing persistence, ingestion, and query performance. ## Description Dowhas an exciting opportunity for aData Engineerlocated inMidland, MI or Houston, TX or Champaign, IL (Dow Delivery Center at UIUC). This position may also supportVirtual Office. This role will make significant technical contributions to critical data initiatives within our team at Dow. You will be responsible for driving the technical implementation and contributing to the design of scalable, Gold-layer data products on the Azure Databricks Lakehouse Platform. This role focuses on solving complex technical challenges, optimization, architecture contribution, and reliability, ensuring our datasets are performant and ready to power advanced use cases, including, + Technical Design Contribution: Collaborate with senior data engineers to translate complex business requirements and ambiguous problem statements into clear, robust, and scalable technical designs and data models (e.g., dimensional modeling, star schemas), and independently drive the implementation of these designs. + Performance Optimization: Design, build, and deploy high-volume data transformation logic using highly optimized PySpark. You will apply advanced techniques to tune Spark jobs and diagnose performance bottlenecks to ensure maximum efficiency and minimal cloud compute cost. + Architecture & Deployment: Contribute significantly to the design and improvement of CI/CD pipelines in Azure DevOps/Git, ensuring reliable, automated, and secure deployment of data solutions across environments. + Diverse Data Integration: Deeply understand and connect to various source systems, demonstrating proficiency in managing data persistence and query performance across diverse technologies like SQL Server, Neo4j, and CosmosDB. + Quality & Governance: Proactively implement and maintain advanced data quality frameworks (e.g., Delta Live Tables, Great Expectations) and monitoring solutions to ensure data reliability for mission-critical applications. + Collaboration & Mentorship: Serve as a go-to technical resource for peers, conducting technical code reviews and informally mentoring Associate Data Engineers on PySpark and Databricks best practices. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Putting the Graph In GraphQL With The Neo4j GraphQL Library](https://www.wearedevelopers.com/videos/257-putting-the-graph-in-graphql-with-the-neo4j-graphql-library) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Cyber Sleuth: Finding Hidden Connections in Cyber Data](https://www.wearedevelopers.com/videos/893-cyber-sleuth-finding-hidden-connections-in-cyber-data) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)