Data Engineer

Insight Global
Redmond, WA, United States
1 day ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Microsoft Azure Cloud Computing Information Systems Continuous Integration Data Architecture Information Engineering Data Governance Data Infrastructure Data Integration Extract Transform Load (ETL) Data Systems
+26 more
Data Warehousing Python (Programming Language) Knowledge Management Machine Learning Microsoft Servers OpenAI Azure Data Lake Software Engineering SQL Databases Unstructured Data Enterprise Search Data Processing Azure Data Factory Retrieval-Augmented Generation Large Language Models Apache Spark Generative AI Indexer Microsoft Fabric Data Lakes AI Platforms Information Technology Data Management Azure Synapse Analytics Data Pipelines Databricks

Job description

Design, develop, and maintain scalable ETL/ELT pipelines that ingest, transform, and serve structured and unstructured data from enterprise systems, documents, communications, and regulatory sources. Build and optimize data architectures, lakehouse solutions, and curated data models to support analytics, machine learning, and generative AI applications. Create document ingestion, metadata extraction, indexing, and retrieval pipelines that enable enterprise search, knowledge management, and RAG-based solutions. Partner with Data Scientists and Applied Scientists to operationalize AI and machine learning solutions through reliable, high-quality data infrastructure. Develop data solutions supporting regulatory intelligence, compliance reporting, requirements extraction, and impact assessment workflows. Implement data quality, governance, lineage, monitoring, and security controls to ensure trusted and auditable enterprise data assets. Optimize data processing performance, scalability, and reliability across cloud-based environments. Collaborate with cross-functional stakeholders to translate business requirements into scalable technical solutions. Drive engineering best practices around testing, CI/CD, observability, documentation, and operational excellence.

Requirements

Bachelor’s degree in Computer Science, Data Engineering, Information Systems, Software Engineering, or related field; equivalent experience considered. 5+ years of experience in Data Engineering or related data platform roles. Advanced proficiency in SQL and Python. Experience building large-scale ETL/ELT pipelines and data integration solutions. Hands-on experience with Azure data technologies including: Azure Databricks Azure Data Factory (ADF) Microsoft Fabric Azure Synapse Analytics Azure Data Lake Storage (ADLS) Experience working with structured, semi-structured, and unstructured data sources. Strong understanding of data modeling, data warehousing, and lakehouse architectures. Experience implementing data governance, data quality, and lineage frameworks. Experience supporting analytics, machine learning, or AI solutions through scalable data infrastructure. Strong communication skills and ability to work across technical and non-technical teams. Experience supporting Generative AI, RAG, LLM, or enterprise search solutions. Experience building document ingestion and knowledge management platforms. Familiarity with vector databases, embeddings, and retrieval architectures. Knowledge of Azure AI Search, Azure OpenAI, or similar AI services. Experience working with legal, compliance, regulatory, governance, or risk-focused organizations. Experience in a large enterprise or Microsoft environment. Experience with Spark, Delta Lake, and modern cloud-native data platforms.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:46 min

Integrating OpenAI REST APIs within Unity environments

Zaid Zaim Zaid Zaim · World Congress 2024

1:03 min

Implementing denormalized schemas with GIN indexes in Postgres

Dharin Shah Dharin Shah · World Congress 2025

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:25 min

Introduction to OpenAI and SingleStore for financial bots

Akmal Chaudhri Akmal Chaudhri · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all