Data Engineer

Jobot
San Diego, CA, United States
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$110,000.0 - $130,000.0
Working hours
Regular working hours
Job source

Tech stack

Data Analysis Microsoft Azure Cyber Security Information Systems Continuous Integration Data Architecture Information Engineering Data Governance Data Integration Extract Transform Load (ETL) Data Transformation Data Security
+22 more
Data Structures Digital Assets Dimensional Modeling Python (Programming Language) Machine Learning Performance Tuning Role-Based Access Control Power BI DataOps SQL Databases Enterprise Data Management Data Ingestion System Availability Apache Spark Microsoft Fabric Pyspark Information Technology SAP S/4HANA Operational Systems Data Management Azure Synapse Analytics Data Pipelines

Job description

We are a leading provider of convenient, high-quality foods designed for busy, health-conscious consumers. For over a century, we’ve focused on delivering products that fit everyday lifestyles, while continually improving our sourcing and production practices to support long-term sustainability and responsible stewardship., We are seeking a Data Engineer to join our team that will be responsible for designing, building, and operating the enterprise data platform that powers analytics, reporting, and advanced data use cases across the organization. This role focuses on data ingestion, transformation, modeling, and platform reliability, with Microsoft Fabric as the primary data engineering and analytics platform.

The Data Engineer will work closely with analytics, IT, and business stakeholders to integrate data from ERP, supply chain, and operational systems into governed, scalable, and high-performance analytical data assets.

This role emphasizes strong engineering discipline, data architecture, and operational excellence rather than ad-hoc analysis or report creation., Design, build, and maintain end-to-end data pipelines using Microsoft Fabric and Azure Synapse, including ingestion, transformation, orchestration, and storage.

  • Develop and manage Fabric Lakehouse and Warehouse architectures to support enterprise analytics and downstream consumption.
  • Implement scalable data ingestion from ERP, supply chain, and operational systems using Fabric Pipelines, Dataflows Gen2, and related tooling.
  • Apply strong data modeling and data engineering best practices, including dimensional modeling, normalization where appropriate, and performance optimization.
  • Create and maintain curated analytical data layers that serve as trusted sources for Power BI semantic models and other consumers.
  • Ensure data quality, reliability, and consistency through validation, monitoring, and structured error handling.
  • Collaborate with analytics and BI teams to support Power BI semantic models, focusing on data structures, performance, and governance rather than report design.
  • Partner with IT security and infrastructure teams to implement role-based access control, data security, and compliance standards.
  • Establish and maintain documentation for data pipelines, schemas, transformations, and architectural decisions.
  • Support deployment and lifecycle management of data assets across development, test, and production environments.
  • Continuously evaluate and adopt new Microsoft Fabric capabilities to improve scalability, performance, and maintainability.

Requirements

  • Bachelor’s degree in computer science, Data Engineering, Information Systems, or a related quantitative field.
  • Strong experience with data engineering concepts, including ETL/ELT, orchestration, data modeling, and pipeline design.
  • Proficiency in SQL for data transformation, validation, and performance tuning.
  • Hands-on experience with Microsoft Fabric and Azure Synapse Analytics.
  • Solid understanding of relational and analytical data modeling techniques.
  • Experience with Python is a must.

Preferred

  • Experience with Microsoft Fabric Lakehouse, Warehouse, Dataflows Gen2, Pipelines, and Notebooks.
  • Understanding of foundational data architecture concepts such as medallion architecture design
  • Familiarity with Power BI semantic models from a data engineering and performance perspective.
  • Experience integrating data from SAP S/4HANA or similar ERP systems.
  • Understanding of supply chain or manufacturing data domains.
  • Exposure to data governance concepts such as lineage, certification, and access control.
  • Experience with automation, CI/CD concepts, or infrastructure-as-code for data platforms.
  • Exposure to machine learning or advanced analytics pipelines is a plus.
  • Experience with Azure DevOps preferred.
  • Experience with Pyspark and exposure to Apache Spark is preferred.
  • Experience with R is preferred.

Benefits & conditions

  • Competitive Salary
  • Full benefits
  • Growth Opportunities
  • Collaborative Culture
  • Stability

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:00 min

Separating dataset creation from low-level software implementation steps

Jan Zawadzki · World Congress 2022

1:24 min

Moving the semantic layer upstream to avoid vendor lock-in

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all