> Markdown version of [/jobs/ext/1839820-databricks-data-lake-engineer](https://www.wearedevelopers.com/jobs/ext/1839820-databricks-data-lake-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Databricks Data Lake Engineer - **Company:** AIT Global, Inc. - **Location:** Jersey City, NJ, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, Big Data, Cloud Database, Cloud Engineering, Continuous Integration, Information Engineering, Extract Transform Load (ETL), Data Migration, Data Warehousing, DevOps, Distributed Computing Environment, Apache Hadoop, Python (Programming Language), Open Database Connectivity, Cloud Services, Azure Machine Learning, SQL Databases, Feature Engineering, Data Ingestion, Azure Data Factory, Informatica Powercenter, Apache Spark, Change Data Capture, Git, Data Lakes, Data Lineage, Apache Kafka, Machine Learning Operations, Restful APIs, Legacy Systems, Databricks - **Published:** July 22, 2026 - **Apply:** https://www.dice.com/job-detail/09e4aadd-4f61-45b9-9333-a5e6380cffbf ## About the Role * 16 years of education with minimum 5+ years of hands-on experience in data engineering with cloud platforms (AWS preferred). * Strong expertise in Databricks, Delta Lake, Apache Spark, and distributed data processing. * Experience with Python, SQL, and ETL/ELT frameworks. * Proven experience with data migration from legacy systems to cloud data lakes. * Deep understanding of data modeling, curation layers, and consumption patterns. * Familiarity with ML workflows, feature engineering, and model operationalization. * Experience with DevOps, Git, CI/CD, and job orchestration tools. * Excellent communication skills with the ability to translate complex concepts into clear business language. Preferred Qualifications: * Experience with AWS Databricks, Azure Data Factory Glue, Airflow, Kafka, Informatica, or similar ingestion tools. * Knowledge of Unity Catalog, Delta Sharing, and enterprise governance frameworks. * Exposure to AI/ML platforms, MLOps, or Databricks Feature Store. * Certifications: Databricks Data Engineer Associate/Professional, AWS and Azure Data Engineer. ## Description The Databricks Data Lake Engineer will design, build, and optimize large scale data pipelines across the full lifecycle of data ingestion, migration, curation, and consumption within a modern lakehouse architecture. This role requires hands on expertise with Databricks, Delta Lake, Spark, and cloud-native data platforms, along with the ability to collaborate effectively with business stakeholders, client, architects, and AI/ML teams., * Data Ingestion - Build scalable ingestion pipelines using Spark, Autoloader, LakeFlow, Informatica, Delta Live Tables, and cloud-native connectors (Kafka, REST, ODBC, CDC - Change Data Capture). * Data Migration - Lead migration of legacy data warehouses, Hadoop clusters, or on prem systems into S3/Delta Lake. * Data Curation - Implement bronze silver gold architecture, enforce quality checks, schema evolution, and governance. * Data Consumption - Deliver curated datasets for BI, analytics, dashboards, and downstream applications. * AI/ML Enablement - Partner with data scientists to prepare feature stores, optimize ML-ready datasets, and support model deployment workflows. * Develop and maintain CI/CD pipelines for Databricks jobs, notebooks, and workflows. * Optimize Spark/SQL/Python jobs for performance, cost efficiency, and reliability. * Implement security, governance, and compliance using Unity Catalog, data lineage, and access controls. * Collaborate with cross-functional teams and communicate technical concepts clearly to non-technical stakeholders. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)