> Markdown version of [/jobs/ext/2393580-data-engineer-materials-discovery-research-institute](https://www.wearedevelopers.com/jobs/ext/2393580-data-engineer-materials-discovery-research-institute). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - Materials Discovery Research Institute - **Company:** Klein Independent School District - **Location:** Skokie, IL, United States - **Experience:** Experienced - **Salary:** $81,456.0 - $112,003.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Data Analysis, Microsoft Azure, Databases, Data Architecture, Data Dictionary, Information Engineering, Data Governance, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Stores, Data Systems, Distributed Computing Environment, Python (Programming Language), Machine Learning, Metadata Repositories, NumPy, Tensorflow, SQL Databases, Workflow Management Systems, Google Cloud, Feature Engineering, Azure Data Factory, Pytorch, Apache Spark, Pandas, Containerization, Core Data, Kubernetes, Information Technology, Data Management, Terraform, Data Pipelines, Software Library, Docker, Databricks - **Published:** August 7, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/87968798/1 ## About the Role * Demonstrated experience owning and evolving data platforms or systems endtoend. * Strong proficiency in SQL and Python, including experience with data analysis and machine learning libraries (e.g., pandas, NumPy, scikitlearn, PyTorch, TensorFlow). * Experience with cloud platforms such as Azure, AWS, or Google Cloud and associated data and analytics services. * Familiarity with infrastructureascode and containerization (e.g., Terraform, Docker, Kubernetes). * Experience with data integration, orchestration tools, and distributed processing frameworks (e.g., Apache Spark, Azure Databricks, Azure Data Factory). * Solid understanding of machine learning fundamentals, feature engineering, evaluation techniques, and experiment reproducibility. * Knowledge of data governance, security, privacy, and compliance best practices. * Strong communication, problemsolving, and technical judgment skills, with the ability to adapt messaging for technical and nontechnical audiences. Professional education and experience requirements for the role include: * Bachelor's degree in Computer Science, Information Technology, Data Science, Engineering, or equivalent combination of education and experience. * Minimum 4 years of experience in a data engineering, analytics engineering, or closely related role. * Demonstrated experience supporting or developing machine learning, statistical About UL Research Institutes and UL Standards & Engagement ## Description We have an exciting opportunity for a Data Engineer at UL Research Institutes, based in our Skokie, Illinois office. This is an onsite opportunity. The Data Engineer role within Materials Discovery focuses on building, maintaining, and supporting reliable data pipelines, data models, and data platforms that enable analytics and machine learning across the institute. The position applies core data engineering practices while contributing selectively to applied data science tasks such as problem definition, data sourcing and preparation, exploratory analysis, and model development. Working closely with data scientists, researchers, senior technical team members, this role plays a key part in onboarding and integrating Engineeringgenerated data into Materials Discovery data infrastructure. The position contributes to architectural and tooling decisions and helps ensure data is wellstructured, accessible, and fit for downstream analytical and modeling workflows. UL Research Institutes: At UL Research Institutes (ULRI), we expand the boundaries of safety science to create a more secure and sustainable world. For more than a century, we have studied the unintended consequences of innovation, designed solutions to mitigate risk and shared our findings with academia, scientists, manufacturers, and policymakers across industries. We identify critical safety and sustainability issues, asking the tough questions because we believe a safer world begins with knowledge. Build a safer, more secure, and sustainable future with us. Join us and work with Materials Discovery teams who conduct the research required to produce that knowledge and put into practice. Materials Discovery Research Institute: The Materials Discovery Research Institute (MDRI) works to develop and deploy new materials with the potential to address current global safety challenges. Pursuing materials that will help produce transformational safety breakthroughs, MDRI harnesses the power of advanced computing and high-throughput experimental methods to create innovative materials that will produce resilience for a sustainable future and protect individual and societal health. We focus on today's critical challenges, working to create new and better materials that will support renewable energy and environmental sustainability. Among our top priorities is research into materials capable of carbon capture and energy storage, with an eye toward reducing the adverse impacts of humanity's reliance upon fossil fuel resources and enabling a transition to renewable energy sources. Above all, our research builds on our commitment to a safer, more sustainable future. What you'll learn and achieve: As the you Data Engineer, will play a key role in the rapid growth of UL as you: * Execute the architecture and technical implementation of MDRI's data platforms, making informed tradeoff decisions related to scalability, performance, cost, security, and reliability. * Define and enforce standards and best practices for data modeling, pipeline design, documentation, data quality, and reproducibility, including implementation of automated data quality checks and validation processes. * Design, build, and evolve data architectures and ETL/ELT pipelines to collect, process, and store data from diverse sources (e.g., laboratory systems, databases, APIs, and external data providers), ensuring data accuracy, completeness, reproducibility, and timeliness. * Evaluate, recommend, and introduce modern data technologies and patterns (e.g., cloudnative services, orchestration frameworks, featureready datasets) aligned with Materials Discovery's current and future needs while proactively addressing system limitations, scaling risks, and performance bottlenecks * Lead integration of disparate data sources into unified, high-quality datasets and ensure data governance, security, and compliance with institutional standards and applicable regulations. * Maintain comprehensive documentation and contribute to data dictionaries and metadata repositories to support longterm sustainability. * Collaborate with researchers and stakeholders to determine effective data and modeling approaches for research, operational, and business challenges. * Assess, select, and justify modeling techniques; perform exploratory data analysis and feature engineering; and develop, train, and evaluate machine learning and statistical models to establish feasibility, baselines, and data requirements. * Clearly document assumptions, inputs, outputs, limitations, and evaluation results, and hand off validated models, feature sets, and documentation for deployment and operationalization. * Act as a technical partner and advisor to researchers, analysts, and leadership on data architecture, analytical feasibility, and strategic trade-offs, while influencing cross-functional technical direction and planning discussions * Assist with troubleshooting complex data and model issues across development and production environments. * Perform other duties as assigned. ## Related Videos - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [Python Data Visualization @ Deepnote (w/ PyViz overview)](https://www.wearedevelopers.com/videos/113-python-data-visualization-deepnote-w-pyviz-overview) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany)