GCP Data Platform Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+2 more
Job description
Experteer Overview In this role you will own the hybrid data platform integration between on-prem Cloudera and Google Cloud. You will establish a repeatable, production-grade pattern for multitenant data architecture starting with GCP as the first cloud tenant. You’ll configure foundational infrastructure, governance, and metadata layers to enable secure cross-environment data access. This is a hands-on, end-to-end engineering position that shapes cross-platform data interoperability and scalability within Capgemini’s data platform initiatives. Compensation / Benefits * Configure and provision the on-prem Cloudera cluster environment, including service accounts and role definitions for multitenant access * Design and implement Apache Iceberg table structures optimized for hybrid on-prem/cloud workloads * Integrate the Iceberg REST Catalog IRC with Apache Gravitino to unify metadata across environments * Establish GCP as a production tenant on the existing on-prem Cloudera cluster and validate the end-to-end architecture * Define and enforce data access governance with RBAC and service account policies across both environments * Lead the full lifecycle from environment setup to production deployment * Publish architecture blueprints and cross-platform interoperability best practices for hybrid data lake implementations Tasks * 5 years of relevant experience in data platform engineering * Hands-on experience with Google Cloud Platform (GCP) * Experience with on-prem Cloudera big data clusters * Experience with Apache Iceberg table format * Experience with Iceberg REST Catalog IRC * Experience with Apache Gravitino for metadata catalog integration * Experience with service account creation and RBAC in hybrid environments Key requirements * Paid time off (vacation, holidays, personal days) * Medical, dental, and vision coverage * Retirement savings plans (401(k) / RRSP) * Life and disability insurance * Employee assistance programs * Local policy-based perks
Requirements
validate the end-to-end architecture * Define and enforce data access governance with RBAC and service account policies across both environments * Lead the full lifecycle from environment setup to production deployment * Publish architecture blueprints and cross-platform interoperability best practices for hybrid data lake implementations Tasks * 5 years of relevant experience in data platform engineering * Hands-on experience with Google Cloud Platform (GCP) * Experience with on-prem Cloudera big data clusters * Experience with Apache Iceberg table format * Experience with Iceberg REST Catalog IRC * Experience with Apache Gravitino for metadata catalog integration * Experience with service account creation and RBAC in hybrid environments Key requirements * Paid time off (vacation, holidays, personal days) * Medical, dental, and vision coverage * Retirement savings plans (401(k) / RRSP) * Life and disability insurance * Employee assistance programs * Local policy-based perks
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Everything a Developer Needs to Know About MCP with Neo4j
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Making Data Warehouses Fast: A Developer’s Story
7 Cloud Computing Trends Coming in 2025 for Developers