> Markdown version of [/jobs/ext/2074573-data-engineer](https://www.wearedevelopers.com/jobs/ext/2074573-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** CVS Health - **Location:** Hartford, CT, United States - **Experience:** Experienced - **Salary:** $79,310.0 - $158,620.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Agile Methodology, Artificial Intelligence, Airflow, Amazon Web Services, Business Analytics Applications, Data Analysis, BigQuery, Cloud Computing, Cloud Engineering, Cloud Storage, Information Systems, Databases, Data as a Services, Data Architecture, Information Engineering, Data Governance, Data Integration, Data Warehousing, DevOps, Distributed Systems, Data Flow Control, Github, Apache Hadoop, Apache Hive, Python (Programming Language), PostgreSQL, Machine Learning, Meta-Data Management, Microsoft SQL Server, MongoDB, NoSQL, Oracle (Applications), PCI Data Security Standards, Power BI, Cloud Services, Standard Sql, PL-SQL, SQL Databases, Data Streaming, Tableau (Software), Teradata SQL, Workflow Management Systems, Enterprise Data Management, Data Logging, Google Cloud, Cloud Platform System, Real Time Systems, Cloud Monitoring, Snowflake, Apache Spark, Indexer, Git, Event Driven Architecture, Data Lakes, Pyspark, Gitlab-ci, Kubernetes, Information Technology, Data Lineage, Collibra, Deployment Automation, Google Cloud Functions, Data Analytics, Qlikview, Star Schema, Apache Kafka, Apache Nifi, Spark Streaming, Data Management, Terraform, Stream Processing, Looker Analytics, Data Pipelines, Serverless Computing, Apache Beam, Amazon Redshift - **Published:** August 15, 2026 - **Apply:** https://dejobs.org/x/x/DD23993D2423479A9AF81B6D610D0310/job/ ## About the Role * 5+ years of professional experience in Data Engineering, Data Platforms, or Data Integration. * 3+ years of hands-on experience with Google Cloud Platform data services. * 3+ years designing, developing, and supporting large-scale batch and streaming data pipelines. * Strong expertise in Python, SQL, Spark/PySpark, Apache Beam, BigQuery, Dataflow, Pub/Sub, and Cloud Composer. * Experience with enterprise data warehouse platforms including BigQuery, Snowflake, Redshift, or Teradata. * Experience with data modeling, data architecture, and distributed systems. * Experience implementing CI/CD pipelines and Infrastructure as Code using Terraform, Git, and cloud-native deployment tools. * Experience working in Agile and SAFe delivery environments. * Strong analytical, troubleshooting, documentation, and problem-solving skills., * Healthcare or Insurance industry experience. * Experience with data governance platforms including Collibra, Privacera, or Unity Catalog. * Experience supporting AI/ML, analytics, and data science workloads. * Experience with MongoDB Atlas and NoSQL data platforms. * Experience building real-time streaming architectures using Kafka and Pub/Sub. * Experience creating Data Flow Diagrams, operational runbooks, and RCA documentation. * Experience with reporting and visualization tools including Power BI, Tableau, QlikView, or Looker. * Experience collaborating across global onshore/offshore engineering teams., Bachelor's degree in Computer Science, Information Systems, Engineering, Data Science, or a related field, or equivalent combination of education and relevant professional experience. ## Description We are seeking a highly skilled Data Engineer to support enterprise data platforms that enable analytics, AI-driven insights, member communications, and business intelligence capabilities. This role requires deep hands-on expertise in Google Cloud Platform (GCP), modern data engineering frameworks, cloud-native architectures, and enterprise data governance., The ideal candidate will have experience delivering scalable data solutions in GCP, building large-scale batch and streaming pipelines, and supporting healthcare or financial services organizations. As a Data Engineer, you will collaborate with architects, product owners, data scientists, and business stakeholders to design, develop, and optimize secure, reliable, and high-performing data platforms in a regulated environment., Data Engineering & Pipeline Development * Design, develop, and maintain scalable batch and real-time data pipelines using GCP services including BigQuery, Dataflow, Pub/Sub, Cloud Composer, Cloud Functions, and Cloud Run. * Build reusable and maintainable ingestion, transformation, and orchestration frameworks using Python, SQL, Spark/PySpark, Apache Beam, . * Develop data integration solutions leveraging APIs, event-driven architectures, and cloud-native services. Cloud Data Platform Development * Build and support enterprise-grade data lake, data warehouse, and streaming solutions across GCP and AWS cloud platforms. * Independently design, develop, and deploy cloud-native data products supporting analytics, reporting, AI/ML, and operational workloads. * Participate in enterprise cloud migration initiatives and modernization efforts, helping define migration strategies and implementation roadmaps. Data Architecture & Modeling * Design and implement logical and physical data models using Star Schema and Snowflake Schema methodologies. * Develop scalable solutions supporting BigQuery, Snowflake, Redshift, Teradata, and other enterprise analytical platforms. * Optimize partitioning, clustering, indexing, and storage strategies to improve performance and cost efficiency. Streaming & Real-Time Processing * Build and maintain high-throughput streaming data solutions using Pub/Sub, Kafka, Dataflow, Spark Structured Streaming, Apache Beam, and Apache NiFi. * Support near real-time data processing requirements across business and operational domains. Data Governance & Compliance * Implement enterprise-grade governance solutions using tools such as Collibra, Privacera, Unity Catalog, and related governance frameworks. * Ensure compliance with HIPAA, GDPR, PCI-DSS, and corporate security standards through robust access controls, lineage tracking, auditing, encryption, and data protection practices. * Support metadata management, data quality monitoring, and data lineage initiatives across the enterprise. Performance, Reliability & Operations * Monitor, troubleshoot, and optimize enterprise data pipelines and cloud infrastructure. * Implement observability using Cloud Monitoring, logging, alerting, and automated recovery solutions. * Create data flow documentation, technical design artifacts, production support procedures, and root cause analyses (RCA). Collaboration & Delivery * Collaborate with business stakeholders, architects, product managers, and engineering teams to deliver scalable data solutions. * Participate in Agile and SAFe delivery frameworks across onshore and offshore teams. * Communicate complex technical concepts effectively to both technical and non-technical stakeholders. Core Technical Skills Cloud Platforms * Google Cloud Platform (GCP) * BigQuery * Dataflow * Pub/Sub * Cloud Composer * Cloud Functions * Cloud Run * Cloud Storage * Cloud Monitoring Data Engineering & Processing * Python * SQL * PySpark * Apache Spark * Apache Beam * Apache Kafka * Apache NiFi * StreamSets * Hadoop Ecosystem Data Warehousing & Analytics * BigQuery * Snowflake * Teradata * Star Schema Design * Snowflake Schema Design Databases * PostgreSQL * SQL Server * Oracle (PL/SQL) * Hive * MongoDB Atlas Workflow Orchestration * Apache Airflow * Cloud Composer DevOps & CI/CD * Git * GitHub Actions * GitLab CI/CD * Terraform ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Tomorrow's cloud data platforms - fully managed database-as-a-service (DBaaS)](https://www.wearedevelopers.com/videos/254-tomorrow-s-cloud-data-platforms-fully-managed-database-as-a-service-dbaas) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk)