> Markdown version of [/jobs/ext/2153456-data-engineer](https://www.wearedevelopers.com/jobs/ext/2153456-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** CVS Health - **Location:** Irving, TX, United States (Remote available) - **Experience:** Experienced - **Salary:** $141,336.0 - $144,200.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Data Analysis, Microsoft Azure, Big Data, Business Software, Code Review, Computer Programming, Databases, Extract Transform Load (ETL), Data Mart, Data Visualization, Dynamical Systems, Apache Hadoop, Hadoop Distributed File System, Python (Programming Language), Power BI, SAS (Software), Software Engineering, SQL Databases, Tableau (Software), Google Cloud, Apache Spark, Backend, Pyspark, Information Technology, Data Pipelines - **Published:** August 20, 2026 - **Apply:** https://dejobs.org/x/x/9C1F6D3DEFCB44359F30904827E11CF8/job/ ## About the Role Master's degree (or foreign equivalent) in Computer Science, Information Technology, Statistics, Analytics, Engineering, or a related field and two (2) years of experience in the job offered or a related occupation. Also requires two (2) years of experience with each of the following: "Big data" platforms including: Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP); SAS or SQL programming languages; Visualization tools, including PowerBI or Tableau; Spark, PySpark, or Scala; Extract/Transform/Load (ETL) processes; Writing application code and deploying to production; and Developing backend services, performing code reviews, and collaborating with peers on software development solutions ## Description Analyze data engineering problems and develop, build and manage large-scale data structures, pipelines and efficient Extract/Load/Transform (ETL) workflows to address complex problems and support business applications. Duties include: develop large scale data structures and pipelines to organize, collect and standardize data to generate insights and addresses reporting needs; write ETL (Extract/Transform/Load) processes, design database systems, and develop tools for real-time and offline analytic processing that improve existing systems and expand capabilities; collaborate with Data Science team to transform data and integrate algorithms and models into automated processes; test and maintain systems and troubleshoot malfunctions; leverage knowledge of Hadoop architecture, HDFS commands, and designing and optimizing queries to build data pipelines; utilize programming skills in Python, Java, or similar languages to build robust data pipelines and dynamic systems; build data marts and data models to support Data Science and other internal customers; integrate data from a variety of sources and ensure adherence to data quality and accessibility standards; analyze current information technology environments to identify and assess critical capabilities and recommend solutions to complex business problems; and experiment with available tools and advise on new tools to provide optimal solutions that meet the requirements dictated by the model/use case. Hybrid position: remote work permitted but must live within commuting distance of designated office location and remain available to report to office as needed and required. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk)