Lead Data Engineer
SDH Systems LLC
San Jose, CA, United States
about 2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source
Tech stack
Artificial Intelligence
Microsoft Azure
Big Data
Cloud Computing
Data Architecture
Information Engineering
Extract Transform Load (ETL)
Data Warehousing
Machine Learning
Performance Tuning
SQL Databases
Data Streaming
+12 more
Azure Service Bus
Feature Engineering
Azure Data Factory
Snowflake
Apache Spark
Data Lakes
Pyspark
Apache Kafka
Azure Synapse Analytics
Stream Analytics
Data Pipelines
Databricks
Job description
- Design, build, and optimize scalable data pipelines using Databricks, Apache Spark, and Azure technologies.
- Architect data warehousing solutions, ensuring seamless integration with cloud platforms and structured/unstructured data sources.
- Collaborate with business stakeholders to understand data needs and develop high-performance analytical solutions.
- Implement ETL/ELT processes leveraging cloud-based technologies such as Azure Data Factory, Snowflake, and Delta Lake.
- Ensure data quality, governance, and security compliance while managing large datasets efficiently.
- Drive performance tuning and optimization for data pipelines, ensuring efficiency across systems.
- Work closely with cross-functional teams to support machine learning and advanced analytics initiatives.
- Provide technical leadership and mentorship to junior data engineers, fostering a culture of innovation and continuous improvement.
- Stay updated on emerging data technologies and recommend strategies to enhance existing architectures.
Requirements
We are seeking a Lead Data Engineer with expertise in Databricks and Data Warehousing to drive data architecture, pipeline development, and optimization efforts. The ideal candidate will play a key role in designing scalable solutions, implementing best practices, and leading data initiatives within a dynamic and collaborative environment., * 8+ years of experience in data engineering, big data processing, and cloud-based solutions.
- Strong expertise in Databricks, Spark (PySpark/SQL), and Delta Lake architecture.
- Proven experience in designing and managing data warehouses using Snowflake, Azure Synapse, or equivalent technologies.
- Deep understanding of data modeling, SQL, and performance optimization.
- Hands-on experience with Azure Data Factory, Event Hubs, and cloud-based ETL processes.
- Solid knowledge of real-time streaming technologies (Kafka, Azure Stream Analytics, or similar).
- Familiarity with ML/AI data pipelines and feature engineering best practices.
- Strong communication and collaboration skills, with experience working in fast-paced, enterprise environments.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on dice.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
BB
Benedikt Bischof
about 4 years ago
DS
Dhannush Subramani
Top Big Data Technologies That You Need to Know
about 4 years ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
LM
Luis Minvielle
How to Become an AI Engineer
over 2 years ago
AJ
Austin Joy
What Are The Top Skills Required For Azure Developers?
over 4 years ago