Senior GCP Data Engineer

EXL SERVICE
United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source

Tech stack

Airflow Data Analysis Computing Platforms Batch Processing BigQuery Cloud Database Cloud Storage Data Architecture Information Engineering Data Governance Extract Transform Load (ETL) Data Structures
+17 more
Data Systems Data Vault Modeling Data Warehousing Dimensional Modeling Distributed Computing Environment Python (Programming Language) Machine Learning Cloudera SQL Databases Feature Engineering Delivery Pipeline Apache Spark Data Lakes Star Schema Data Management Data Pipelines Automation Anywhere

Job description

Job Description: We are seeking a Senior Data Engineer (Lead) to drive and own end-to-end data engineering initiatives. This role will lead all data engineering efforts, working closely with Data Science and Analytics teams to design scalable data platforms, enable advanced analytics, and support machine learning use cases.

The ideal candidate will bring deep expertise in cloud data engineering (GCP) , strong data modeling capabilities , and proven experience in leading enterprise-grade data solutions.

  • Responsibilities: Lead and manage end-to-end Data Engineering delivery across projects and initiatives
  • Act as the primary technical owner for data pipelines, architecture, and platform design
  • Mentor and guide a team of data engineers, ensuring best practices and coding standards
  • Design, build, and optimize scalable data pipelines on GCP
  • Define and implement modern data architectures (data lake, lakehouse, warehouse)
  • Ensure high performance, reliability, and data quality across pipelines
  • Partner closely with Data Science teams to enable ML/AI workflows
  • Translate business and modeling requirements into optimized data structures
  • Support feature engineering, model training, and deployment pipelines
  • Design logical and physical data models for analytics and ML use cases
  • Implement dimensional modeling (Star/Snowflake schemas) and data vault where applicable
  • Optimize datasets for performance, scalability, and usability
  • Build and manage solutions using GCP services such as: BigQuery, Cloud Composer (Airflow), Cloud Storage, Dataproc
  • Ensure security, governance, and cost optimization on GCP

Requirements

Do you have experience in Spark implementation?, * 8+ years of experience in Data Engineering , with leadership experience

  • Strong expertise in GCP ecosystem and services
  • Proficiency in SQL, Python , and/or Scala
  • Hands-on experience with ETL frameworks and distributed processing
  • Solid experience in Dimensional modeling, Data warehousing concepts, Data structures for ML/analytics
  • Experience with Apache Spark and Real-time and batch processing frameworks
  • Experience working with cross-functional teams (Data Science, Analytics, Business)
  • Proven ability to lead, mentor, and drive delivery
  • Strong ownership mindset with leadership capabilities
  • Excellent problem-solving and architectural thinking
  • Ability to operate in a fast-paced, collaborative environment

Preferred Qualifications:

  • Experience in ML data pipelines / feature stores
  • Knowledge of data governance, lineage, and quality frameworks
  • Exposure to healthcare/payor domain (nice to have)
  • Certifications in GCP (Professional Data Engineer)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:27 min

Explaining query execution overhead and caching limitations in BigQuery

Adnan Rahic · JS Congress

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:30 min

Leveraging BigQuery ML for scalable SQL-based segmentation experiments

Julian Joseph · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

Videos

See all

Related articles

See all