> Markdown version of [/jobs/ext/2872693-data-engineer-gcp-aws-databricks](https://www.wearedevelopers.com/jobs/ext/2872693-data-engineer-gcp-aws-databricks). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer GCP AWS & Databricks - **Company:** iShare Inc - **Location:** Ontario, CA, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Agile Methodology, Airflow, Amazon Web Services, Amazon S3, Big Data, BigQuery, Cloud Database, Cloud Storage, Computer Programming, Continuous Integration, Data Architecture, Information Engineering, Data Integration, Extract Transform Load (ETL), Data Security, Data Warehousing, Data Flow Control, Python (Programming Language), Cloud Services, Cloudera, SQL Databases, Unstructured Data, Workflow Management Systems, Google Cloud, Apache Spark, Git, Data Lakes, Pyspark, Infrastructure Automation Frameworks, Real Time Data, Apache Kafka, Data Management, Video Streaming, Terraform, Data Pipelines, Databricks - **Published:** September 13, 2026 - **Apply:** https://www.wayup.com/i-j-Data-Engineer-GCP-AWS-Databricks-iShare-Inc-656812608462550/ ## About the Role 5+ years of professional data engineering experience. Strong hands-on experience with both GCP and AWS. Expertise in Databricks, Apache Spark, and PySpark. Strong programming skills in Python and SQL. Proven experience developing ETL/ELT pipelines and cloud data platforms. Experience with data warehouses, data lakes, and dimensional data modeling. Experience with orchestration tools such as Apache Airflow or Google Cloud Composer. Understanding of data security, governance, monitoring, and quality frameworks. Strong analytical, troubleshooting, and problem-solving skills. Excellent communication and cross-functional collaboration skills. Preferred Qualifications Experience with GCP services such as BigQuery, Cloud Storage, Dataflow, Dataproc, Pub/Sub, and Cloud Composer. Experience with AWS services such as S3, Glue, EMR, Redshift, Lambda, and Kinesis. Familiarity with Delta Lake and Databricks Lakehouse architecture. Experience with streaming technologies such as Apache Kafka. Familiarity with CI/CD, Git, Terraform, and cloud infrastructure automation. Experience working in Agile development environments. Must-Have Skills GCP | AWS | Databricks | Apache Spark | PySpark | Python | SQL | ETL/ELT | Airflow/Cloud Composer | Data Warehousing | Data Lakes ## Description We are seeking an experienced Data Engineer to design, develop, and maintain scalable cloud-based data pipelines and data platforms. The ideal candidate will have strong hands-on experience with GCP, AWS, Databricks, Apache Spark, PySpark, Python, and SQL. This role requires expertise in building reliable ETL/ELT pipelines, integrating data from multiple sources, and optimizing cloud data solutions for performance, security, scalability, and cost., Design, develop, and maintain scalable ETL/ELT data pipelines. Build cloud-based data solutions using GCP and AWS services. Use Databricks, Apache Spark, and PySpark for large-scale data processing. Integrate structured, semi-structured, and unstructured data from multiple sources. Develop and optimize batch and real-time data-processing workflows. Improve pipeline performance, reliability, scalability, and cost efficiency. Implement data-quality checks, monitoring, security, and governance standards. Design and support cloud data warehouses and data lakes. Troubleshoot production issues and perform root-cause analysis. Collaborate with data architects, analysts, application teams, and business stakeholders. Create and maintain technical documentation for pipelines, data models, and workflows. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)