> Markdown version of [/jobs/ext/3288865-data-engineer](https://www.wearedevelopers.com/jobs/ext/3288865-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Southern Hobby Distribution - **Location:** United States (Remote available) - **Experience:** Experienced - **Contract:** Temporary contract - **Skills:** Sql Data Warehouse, Artificial Intelligence, Amazon Web Services, Amazon S3, Data Analysis, ARM Architecture, User Authentication, Software as a Service, Code Review, Continuous Integration, Data as a Services, Data Architecture, Data Dictionary, Information Engineering, Data Integration, Extract Transform Load (ETL), Data Mart, Software Debugging, Dimensional Modeling, Github, Identity and Access Management, Job Scheduling, Python (Programming Language), Power BI, Shopify, Qualtrics, File Transfer Protocol (FTP), Sql Optimization, Large Language Models, Snowflake, Boto3, AI Coding Agents, Rate Limiting, Git, Pandas, Build Management, Low-code, Integration Frameworks, Claude, Virtual Agents, Hubspot, Restful APIs, Pagination, Data Pipelines, Jenkins - **Published:** September 28, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=cb16d38559fc0b69 ## About the Role * 3-5 years of hands-on data engineering experience. * Hands-on experience building star-schema data marts (fact/dimension design, grain definition, surrogate keys, and SCD handling) for BI consumption. * Solid understanding of dimensional modeling and warehouse layering concepts. * Strong Python (pandas, requests, boto3, snowflake-connector) for production pipelines. * Advanced SQL and working experience with Snowflake or a comparable cloud data warehouse. * Experience integrating with REST APIs, including authentication, pagination, and rate limiting. * Familiarity with AWS services used in data workflows (S3, Lambda, Secrets Manager, IAM). * Experience with CI/CD and job scheduling (Jenkins, GitHub Actions, or similar). * Experience with Snowflake Cortex, semantic models, or LLM/AI-assisted data workflows. * Demonstrated ability to work effectively with AI coding assistants, including strong prompting, critical review of generated code, and sound judgment on when not to rely on them. Nice to Have * Experience with industry-leading ETL tools, n8n, or similar low-code integration tools. * dbt or an equivalent transformation and testing framework. * Power BI data modeling or DAX exposure. * Retail, wholesale, or consumer-products domain experience (ERP, CRM, PLM, e-commerce data). Success in the First 6-12 Months * Independently own eight to ten production pipelines end to end. * Reduce recurring pipeline failures through improved monitoring and data quality checks. * Contribute reusable pipeline templates that shorten onboarding time for new sources. #LI-Remote ## Description Collectivus is building a modern data platform on Snowflake to serve its portfolio of brands. As a Data Engineer, you will design, build, and maintain the pipelines that move data from our source systems (ERP, CRM, e-commerce, marketing, survey, and finance platforms) into Snowflake, and shape it into reliable, well-modeled data for analytics, Power BI reporting, and AI-driven applications. You will work within an established architecture and set of engineering standards, partnering with the Data Architect on design decisions and with the Data Analysis team on delivery. Our data engineering team actively uses AI assistants in day-to-day development, and we expect engineers to adopt these tools thoughtfully and responsibly., Pipeline Development & Data Integration * Build and maintain Python-based ETL/ELT pipelines that ingest data from REST APIs, SFTP/flat files, and SaaS platforms (e.g., Qualtrics, Shopify, UKG, ERP, HubSpot) into Snowflake. * Implement incremental loads, MERGE-based upserts, watermark/control-table patterns, soft deletes, and idempotent reruns following team conventions. Data Modeling & Warehousing * Develop and tune Snowflake SQL across raw, staged, and presentation layers, including views, tasks, and streams. * Design and build data marts in Snowflake (fact and dimension tables, conformed dimensions, and aggregate layers) that serve Power BI reporting and downstream analytics. Orchestration, Security & Data Quality * Schedule and orchestrate jobs using GitHub Actions, monitor runs, and resolve failures. * Manage secrets and credentials securely via AWS Secrets Manager and Snowflake key-pair authentication. * Build data quality checks and assertion tests, and detect and remediate schema drift from upstream sources. AI & Advanced Data Services * Support Snowflake Cortex and MCP-based data services (semantic views, search services) by preparing and governing the underlying data. * Use AI coding assistants (e.g., Claude) as part of the daily workflow to accelerate pipeline development, code review, debugging, and documentation, while retaining full ownership of correctness and quality. Collaboration, Documentation & Engineering Standards * Collaborate with analysts and business stakeholders to translate reporting needs into data models and Power BI-ready datasets. * Write clear runbooks and data dictionaries. * Participate in code reviews and follow Git branching and pull request standards. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)