GCP Data Lead

HCLTech
West Sussex, UK
10 days ago
Apply on uk.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Adaptable Database Systems Airflow Amazon Web Services Application Frameworks Automation of Tests Microsoft Azure Big Data BigQuery Cloud Database Computer Programming Databases
+28 more
Continuous Integration Data Governance Extract Transform Load (ETL) Dataspaces Data Systems Shard (Database Architecture) IBM InfoSphere DataStage DevOps Distributed Data Store Distributed Systems Apache HBase Identity and Access Management Python (Programming Language) Enterprise Messaging Systems MongoDB NoSQL Performance Tuning SQL Databases Data Streaming Data Ingestion Snowflake Apache Spark Containerization Kubernetes Apache Flink Apache Kafka Data Pipelines Amazon Redshift

Job description

The IDEA Lab is expanding its GCP Data Products capability and is looking for a highly skilled Data Engineer to build, optimise, and scale data pipelines and large-scale data processing workloads in a cloud-native environment. You will work with modern distributed data systems, contribute to data platform modernisation, and support large-scale ingestion, transformation, and analytics workloads across the Business Transaction Banking platform., Design and deliver end-to-end data pipelines on cloud platforms (GCP preferred). Build scalable data ingestion, transformation, and processing workflows using distributed technologies such as Spark, Flink, Storm, or similar. Develop robust ELT/ETL processing and migration pipelines, including support for legacy Datastage decommissioning and modernisation. Work with a variety of database technologies including relational, NoSQL, MPP and columnar stores (BigQuery, Redshift, Azure SQLDW, HBase, MongoDB). Implement streaming and messaging-based pipelines using Kafka, Pulsar or Pub/Sub. Build optimised, scalable data models to support diverse consumption patterns, applying partitioning, sharding, bucketing and aggregation strategies. Apply performance tuning and optimisation across storage, compute and query layers. Ensure secure handling of data including authentication, authorisation, encryption in transit/at rest, and cloud-native security controls. Implement monitoring, alerting and observability for large-scale distributed data workloads. Use orchestration tools such as Cloud Composer, Airflow or equivalent to operationalise pipelines. Contribute to CI/CD, containerisation, Kubernetes-based deployments, and automated testing practices. Participate in data governance, metadata, catalogue and lineage processes as needed. Collaborate with engineers, architects and SMEs to deliver stable, high-quality data products.

Requirements

Strong programming skills in Java (preferred) , Python , or Scala . Hands-on experience with cloud data services (GCP preferred; Azure/AWS acceptable). Practical experience with distributed data processing frameworks such as Spark (Core/SQL/Streaming), Flink, or Storm. Experience across multiple database technologies-Relational, NoSQL, MPP, columnar. Strong knowledge of data ingestion, transformation and messaging systems : Kafka, Pulsar, Pub/Sub, etc. Understanding of designing scalable data models for varied access patterns. Experience with performance tuning , cost-optimisation and scaling strategies. Experience delivering large-scale big data solutions in batch and/or streaming environments, on cloud or on-premise. Good familiarity with the wider data ecosystem and open-source frameworks. Experience with orchestration (Airflow/Composer) and workflow automation. Understanding of DevOps for data systems: CI/CD, containers, Kubernetes and automated testing. Knowledge of security for big-data systems including IAM, encryption, and cluster-level controls. Basic understanding of monitoring and alerting for distributed systems. Knowledge of dimensional modelling (star, snowflake, normalized/denormalized). Awareness of data governance, cataloguing and lineage tools.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on uk.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all