Google Cloud Platform Data Engineer

Cosourcing Partners LLC
Chicago, IL, United States
2 months ago

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

BigQuery Cloud Computing Cloud Database Cloud Storage Software Quality Continuous Integration Data Architecture Information Engineering Data Infrastructure Extract Transform Load (ETL) Data Transformation Data Security
+30 more
Data Warehousing DevOps Fault Tolerance Data Flow Control Github Identity and Access Management Python (Programming Language) Machine Learning Performance Tuning SQL Databases Data Streaming Talend Unstructured Data Data Logging Data Processing Freeform SQL Google Cloud Feature Engineering Data Build Tool (dbt) Apache Spark Build Server Git Information Technology Deployment Automation Data Management Data Lakehouse Software Version Control Data Pipelines Apache Beam Jenkins

Job description

We are seeking a highly skilled Google Cloud Platform Data Engineer to design, develop, and optimize scalable data solutions on Google Cloud Platform (Google Cloud Platform). The ideal candidate will have strong expertise in building robust batch and streaming pipelines, implementing modern data architectures, and enabling reliable, high-quality data platforms for analytics, reporting, and machine learning use cases., Data Engineering & Pipeline Development

  • Design, build, and optimize scalable batch and real-time (streaming) data pipelines using Google Cloud Platformnative services.
  • Develop and maintain data ingestion frameworks leveraging tools such as Pub/Sub, Dataflow, and Cloud Storage.
  • Implement data transformation pipelines using BigQuery, dbt, and Python-based workflows.
  • Ensure efficient handling of large-scale structured and unstructured datasets. Data Modeling & Architecture
  • Design and implement high-performance data models for cloud-based data lakes, data warehouses, and analytics platforms.
  • Optimize data schemas and partitioning strategies in BigQuery for performance and cost efficiency.
  • Support modern architectures such as medallion (bronze/silver/gold) layers and lakehouse patterns.

Development & Coding

  • Write advanced SQL queries for transformation, validation, and analytics.
  • Develop scalable data processing logic using Python and/or Apache Beam.
  • Build reusable, modular, and maintainable code for data workflows.

Data Quality, Observability & Reliability

  • Implement and maintain data quality checks, validation rules, and anomaly detection frameworks.
  • Enable data observability through monitoring, logging, and alerting mechanisms.
  • Ensure highly reliable data pipelines with fault tolerance and error handling strategies.

ETL/ELT Modernization

  • Support migration and modernization efforts from legacy ETL tools (e.g., Talend) to Google Cloud Platform-native ELT frameworks (dbt).
  • Optimize existing pipelines for performance, scalability, and maintainability in cloud environments.
  • Drive adoption of ELT best practices using BigQuery as the compute engine.

Collaboration & Stakeholder Engagement

  • Collaborate with data architects, business analysts, and machine learning teams to deliver trusted datasets.
  • Translate business requirements into scalable data solutions.
  • Provide technical guidance and support for downstream analytics and reporting use cases.

Best Practices & Governance

  • Drive adoption of best practices in cloud data engineering, CI/CD, and DevOps.
  • Implement secure data access controls using IAM roles, policies, and governance frameworks.
  • Follow standards for code quality, version control (Git), and automated deployments.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or related field.
  • 4+ years of experience in data engineering or data platform development.
  • Hands-on experience with Google Cloud Platform (Google Cloud Platform) services:

  • BigQuery
  • Dataflow
  • Pub/Sub
  • Cloud Storage

  • Strong proficiency in SQL and Python.
  • Experience with dbt (Data Build Tool) or similar ELT frameworks.
  • Experience building batch and streaming data pipelines.

Preferred Skills

  • Experience with Apache Beam or Spark.
  • Familiarity with Talend or other ETL tools and migration to cloud-native solutions.
  • Knowledge of data lakehouse architectures and modern data stack.
  • Experience with CI/CD tools (e.g., GitHub Actions, Cloud Build, Jenkins).
  • Understanding of data security, governance, and compliance standards.
  • Exposure to machine learning data pipelines and feature engineering.

Key Competencies

  • Strong problem-solving and analytical skills
  • Ability to work in cross-functional teams
  • Excellent communication and documentation skills
  • Focus on performance optimization and scalability
  • Attention to data quality and reliability

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

Videos

See all

Related articles

See all