> Markdown version of [/jobs/ext/2028512-data-engineer](https://www.wearedevelopers.com/jobs/ext/2028512-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Lhi Group Ltd - **Location:** Woodbridge Township, NJ, United States - **Experience:** Starter - **Salary:** $88,000.0 - $105,000.0 - **Contract:** Internship / Graduate position - **Skills:** Amazon Web Services, Data Analysis, Microsoft Azure, Cloud Computing, Data Architecture, Data Cleansing, Data Deduplication, Information Engineering, Data Transformation, Data Structures, Data Systems, Data Visualization, Relational Databases, Python (Programming Language), Object-Oriented Software Development, Raw Data, Cloud Services, SQL Databases, Google Cloud, Apache Spark, Git, Data Lakes, Pyspark, Semi-structured Data, Core Data, Information Technology, Software Version Control, Data Pipelines, Databricks - **Published:** August 11, 2026 - **Apply:** https://www.dice.com/job-detail/bd378afe-f3f4-4008-b993-01084fb47768 ## About the Role This opportunity is well suited to a Data Engineer, Data Analyst, or recent graduate with strong foundational skills in SQL and Python who is looking to build hands-on experience with Databricks, Apache Spark, PySpark, and modern lakehouse architecture., * 0-2 years of experience in Data Engineering, Analytics, Computer Science, or a related field; internships and academic projects are considered * Foundational knowledge of SQL, including joins, filtering, aggregations, and querying relational data * Experience with Python and an understanding of basic object-oriented programming concepts * Understanding of core data concepts, including tables, schemas, relational databases, and data structures * Familiarity with Git or demonstrated ability to quickly learn version-control practices * Strong interest in developing expertise across Databricks, Spark, and cloud technologies * Strong analytical, problem-solving, and communication skills Preferred Qualifications * Exposure to Databricks, Apache Spark, or PySpark * Understanding of Delta Lake or medallion architecture * Experience with Azure, AWS, or Google Cloud Platform * Databricks certification * Exposure to BI, reporting, or data visualization tools * Academic or project experience building data pipelines or working with data transformation workflows ## Description * Support the development and maintenance of data pipelines across Bronze, Silver, and Gold medallion architecture layers * Ingest and organize raw data within the Bronze layer * Support data cleansing, deduplication, validation, and schema enforcement within the Silver layer * Develop SQL and PySpark transformations based on defined business and technical requirements * Assist with data quality, testing, monitoring, and troubleshooting of pipeline jobs * Work with structured and semi-structured data within cloud-based data environments * Collaborate closely with senior data engineers to learn and apply modern data engineering practices * Contribute to scalable and reliable data solutions using version-controlled development practices, This is an excellent opportunity for an early-career professional looking to develop a strong foundation in data engineering, cloud data platforms, Databricks, Spark, and lakehouse architecture while working alongside experienced engineers. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Governance in the Era of AI](https://www.wearedevelopers.com/videos/1622-data-governance-in-the-era-of-ai) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)