> Markdown version of [/jobs/ext/797398-data-engineer](https://www.wearedevelopers.com/jobs/ext/797398-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** CVS Health - **Location:** North Haven, CT, United States (Remote available) - **Salary:** $142,418.0 - $158,620.0 - **Contract:** Internship / Graduate position - **Skills:** Java (Programming Language), Agile Methodology, Artificial Intelligence, Amazon Web Services, Data Analysis, Microsoft Azure, Big Data, Business Software, Cloud Computing, Computer Programming, Databases, Continuous Integration, Data Architecture, Extract Transform Load (ETL), Data Mart, Data Visualization, DevOps, Distributed Systems, Dynamical Systems, R (Programming Language), Apache Hadoop, Hadoop Distributed File System, Python (Programming Language), Machine Learning, MySQL, NoSQL, Power BI, SAS (Software), Scala (Programming Language), Software Engineering, SQL Databases, Tableau (Software), Pytorch, Apache Spark, Git, Pyspark, Scikit Learn, Information Technology, Machine Learning Operations, Data Pipelines, Jenkins - **Published:** June 22, 2026 - **Apply:** https://dejobs.org/x/x/CE554E36A63F4C51AB0E8558AE108F48/job/ ## About the Role Requirements: Master's degree (or foreign equivalent) in Computer Science, Data Science, Statistics, Mathematics, Analytics, or a related field. Requires completion of a university-level course, research project, internship, or thesis in each of the following: * CI/CD, Jenkins, GIT, or DevOps; * Programming in Java, Python, and R; * SAS or SQL programming languages; * Agile methodologies or SAFe Software Development Principles; * Spark, PySpark, and Scala; * Hadoop architecture or HDFS commands; * MySQL and NoSQL; * Visualization tools: PowerBI and Tableau; * Designing and optimizing queries to build data pipelines; * Machine, learning, statistical analysis, and predictive modeling; * NLP (Scikit, SpaCity, Pytorch, or Spark NLP); * Vertex-AI; * PyData ecosystem; * Writing Extract/Transform/Load (ETL) processes; * Designing data architectures, including data pipelines, distributed computing engines, and machine learning infrastructure design; and * Developing and deploying predictive models or ML systems in a cloud environment (GCP, AWS, or Azure). ## Description Position Summary: Aetna Resources, LLC, a CVS Health company, is hiring for the following role in Hartford, CT: Data Engineer to develop, build, and manage large-scale data structures, pipelines, and efficient Extract/Load/Transform (ETL) workflows to address complex problems and support business applications. Duties include: develop large scale data structures and pipelines to organize, collect and standardize data to generate insights and addresses reporting needs; write ETL (Extract/Transform/Load) processes, design database systems, and develop tools for real-time and offline analytic processing that improve existing systems and expand capabilities; collaborate with Data Science team to transform data and integrate algorithms and models into automated processes; test and maintain systems and troubleshoot malfunctions; leverage knowledge of Hadoop architecture, HDFS commands, and designing and optimizing queries to build data pipelines; utilize programming skills in Python, Java, or similar languages to build robust data pipelines and dynamic systems; build data marts and data models to support Data Science and other internal customers; integrate data from a variety of sources and ensure adherence to data quality and accessibility standards; analyze current information technology environments to identify and assess critical capabilities and recommend solutions to complex business problems; and experiment with available tools and advise on new tools to provide optimal solutions that meet the requirements dictated by the model/use case. Telecommuting available. Multiple positions. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Branch your database like your code: How schema changes and pull requests go hand in hand](https://www.wearedevelopers.com/videos/350-branch-your-database-like-your-code-how-schema-changes-and-pull-requests-go-hand-in-hand) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)