Senior Data Engineer

CVS Health
New York, NY, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$83,430.0 - $222,480.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Big Data BigQuery Cloud Computing Data Architecture Data Governance Data Integration Extract Transform Load (ETL) Data Mart Data Systems Data Warehousing
+12 more
Database Queries Dimensional Modeling Python (Programming Language) Machine Learning Performance Tuning DataOps SQL Databases Google Cloud Data Storage Technologies Build Management Pyspark Data Pipelines

Job description

We’re seeking a Sr. Data Engineer to design and implement data pipelines that power analytical capabilities. This hands-on role requires an understanding of data engineering best practices and the ability to translate business requirements into technical solutions., + Data Pipeline Development: Design and build ETL/ELT data pipelines to ingest, process, and transform large datasets from multiple sources.

  • Performance Optimization: Implement best practices for performance tuning, partitioning, and clustering to optimize data queries and reduce costs.

  • Data Quality & Governance: Establish and enforce data quality standards, data governance frameworks, and security policies for data storage and access.

  • Data Modeling & Architecture: Develop and optimize data models and schemas to support analytics, reporting, and machine learning requirements.

  • Data Integration & Transformation: Collaborate with data scientists and analysts to design data solutions that integrate with BI tools and machine learning models.

  • Documentation & Knowledge Sharing: Create comprehensive documentation for data pipelines, workflows, and processes. Share best practices and mentor junior data engineers.

Requirements

  • 5+ years of applicable work experience

  • Proficiency in Python, specifically with ETL pipelines.

  • Strong proficiency in SQL and experience in developing complex queries.

  • Familiarity with pySpark, DBT, or other similar frameworks.

  • Experience deploying data pipelines in a cloud environment (Azure, AWS, GCP).

  • Understanding of data warehousing concepts, dimensional modeling, and building data marts.

  • Excellent communication and interpersonal skills, with the ability to collaborate effectively with data scientists, analysts, and product owners.

Preferred Qualifications

  • Knowledge of data governance best practices in a cloud environment.

  • Experience with data design in BigQuery

  • Experience working with the Epic data model.

  • Experience working with healthcare data (Claims and Admissions)

Education and Experience

  • College degree or certification in related fields

Benefits & conditions

This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls. The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors. This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above.

Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong.

Great benefits for great people

We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families.

This full-time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well-being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:00 min

Separating dataset creation from low-level software implementation steps

Jan Zawadzki · WWC 2022

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all