Data Engineer

Databricks
United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Microsoft Azure Business Intelligence Development Unix Cloud Computing Databases Data as a Services Data Integration Extract Transform Load (ETL) Data Systems IBM DB2 Healthcare Effectiveness Data and Information Set Python (Programming Language)
+12 more
Oracle (Applications) Power BI Cloudera SQL Databases Teradata SQL Workflow Management Systems Google Cloud Apache Spark Pyspark Tools for Reporting Data Pipelines Databricks

Requirements

3 years of experience with data pipeline and workflow management tools (e.g., Apache Spark(Dataproc), Google Cloud Platform Tools, Databricks, PySpark).

3 Years of experience building ETL and Data Integration pipelines in Python, PySpark and knowledge of Informatica is a plus

3 years of experience in a reporting tool like Power BI & Tableau. (Power BI preferred)

3 years of experience working with On Prem databases like Oracle, Teradata and DB2

3 years of experience with Cloud platforms (Google Cloud Platform and Azure) and their respective data services

3 years of deep SQL experience & Unix experience

3 years of experience working with a variety of technology systems, designing solutions or developing data solutions in healthcare. Hedis experience is a big plus.

Educational Qualifications: -

  • Engineering Degree BE/ME/BTech/MTech/BSc/MSc.

Technical certification in multiple technologies is desirable.

Skills: -

Mandatory skills

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann +2 · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:04 min

Defining timestamps and the international standard format

Denny Biasiolli Denny Biasiolli · Europe 2026 Virtual

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all