GCP ETL Engineer

OpenKyber LLC
United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Data Analysis Big Data BigQuery Data Integration Extract Transform Load (ETL) Data Transformation DevOps Apache Hive Performance Tuning Query Optimization SQL Stored Procedures SQL Databases
+6 more
Enterprise Data Management Data Ingestion Apache Spark Caching Data Lakes Real Time Data

Job description

Design, develop, and optimize large-scale data processing pipelines using Apache Spark (Spark SQL, Spark Core, and DataFrames) for high-volume batch and real-time data workloads.

Write complex and high-performance SQL queries, stored procedures, and data transformations to support data ingestion, cleansing, aggregation, and reporting requirements.

Perform advanced performance tuning and troubleshooting of Spark jobs and SQL workloads, including partitioning, joins, caching, and query optimization for improved efficiency.

Build and maintain ETL/ELT workflows by integrating data from multiple structured and unstructured sources into enterprise data platforms and data lakes.

Work with BigQuery for data analysis, query optimization, and large-scale analytical reporting; knowledge of schema design and cost-efficient query execution is an added advantage.

Collaborate with cross-functional teams including Data Architects, Analysts, and DevOps teams to ensure scalable, secure, and reliable big data solutions aligned with business requirements.

For applications and inquiries, contact: hirings@openkyber.com

Requirements

Do you have experience in Spark implementation?

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:15 min

Reversing the caching model for artifact delivery

Thijs Feryn Thijs Feryn · WWC Europe 2026

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all