> Markdown version of [/jobs/ext/2615519-data-engineer](https://www.wearedevelopers.com/jobs/ext/2615519-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** AllSTEM Connections - **Location:** Ontario, CA, United States - **Experience:** Expert - **Contract:** Temporary contract - **Skills:** Agile Methodology, Artificial Intelligence, Airflow, Data Analysis, Unit Testing, Big Data, BigQuery, Code Review, Information Systems, Data Validation, Information Engineering, Extract Transform Load (ETL), Data Systems, Dimensional Modeling, Distributed Computing Environment, Python (Programming Language), Operational Databases, Scrum Methodology, Cloud Services, Cloudera, Data Streaming, Technical Data Management Systems, Enterprise Data Management, Data Processing, Google Cloud, Warehouse Management Systems, Sql Optimization, Data Build Tool (dbt), Reliability of Systems, Git, Pyspark, Information Technology, Apache Kafka, Cloud Optimization, Software Version Control, Data Pipelines - **Published:** August 2, 2026 - **Apply:** https://www.dice.com/job-detail/28629468-130d-46d1-8c95-c9b1579da96c ## About the Role Education: Bachelor's degree in Computer Science, Engineering, Information Systems, or a related quantitative field, or equivalent practical experience. Experience Baseline: 4+ years of professional experience in Data Engineering. Technical Mastery: oHands-on experience building production data solutions on Google Cloud Platform (Google Cloud Platform). oStrong proficiency with Dataproc, BigQuery, advanced SQL, and dbt (Data Build Tool). oDeep understanding of data modeling techniques, dimensional modeling, and analytical data warehouse architecture. oProven experience building scalable ETL/ELT pipelines for large datasets. oWorking knowledge of Git-based version control and CI/CD best practices. Core Competencies: Exceptional analytical and troubleshooting abilities; strong communication skills; ability to work independently within fast-paced Agile teams. Preferred Attributes Experience with workflow orchestration tools such as Apache Airflow. Familiarity with event streaming platforms (e.g., Apache Kafka) and distributed data processing using PySpark and Python. Domain expertise within retail, apparel, supply chain platforms, transportation logistics, or Warehouse Management Systems (WMS). ## Description We are seeking a skilled and driven Data Engineer to join our enterprise data and analytics organization. In this role, you will focus on developing and supporting scalable data products and robust ETL/ELT pipelines that power advanced analytics and AI initiatives across sourcing, transportation, and warehouse management domains. This is a hands-on technical position where you will collaborate closely with product managers, data architects, business analysts, and fellow engineers to design, build, test, and maintain cloud-native data solutions. Utilizing modern Google Cloud Platform (Google Cloud Platform) technologies, you will transform enterprise data into trusted, high-quality assets that drive operational excellence. If you have a strong background in building scalable data pipelines on Google Cloud Platform, we want to hear from you., Data Pipeline & Infrastructure Development ETL/ELT Architecture: Design, develop, test, and maintain scalable ETL/ELT data pipelines on Google Cloud Platform (Google Cloud Platform) to ingest, transform, validate, and publish multi-source enterprise data. Cloud Optimization: Build and optimize high-performance data solutions leveraging Dataproc, BigQuery, advanced SQL, and dbt. Data Modeling: Develop and maintain rigorous data models (including dimensional modeling and analytical warehouse concepts) supporting reporting and AI use cases. Collaboration & Engineering Standards Cross-Functional Delivery: Partner with product managers, business analysts, and architects to translate business requirements into effective technical data solutions. Performance & Cost Efficiency: Continuously optimize data processing performance, system reliability, and cloud resource cost-efficiency. Quality Assurance: Perform unit testing, automated data validation, and production support to ensure flawless data quality and operational stability. Agile Delivery & Documentation Agile Ceremonies: Participate actively in sprint planning, backlog refinement, code reviews, and cross-functional team ceremonies. Technical Documentation: Document technical designs, data flows, and implementation details to ensure maintainability and knowledge sharing., AllSTEM Connections participates in the E-Verify program in certain locations as required by law. Learn more about the E-Verify program. _Participation_Poster_ES.pdf We also consider for employment qualified applicants regardless of criminal histories, consistent with legal requirements, including, if applicable, the City of Los Angeles' Fair Chance Initiative for Hiring Ordinance. Pursuant to applicable state and municipal Fair Chance Laws and Ordinances, we will consider for employment-qualified applicants with arrest and conviction records, including, if applicable, the San Francisco Fair Chance Ordinance. For Los Angeles, CA applicants: Qualified applications with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)