Data Engineer

EXL LIFE PRO
United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$110,000.0 - $130,000.0
Working hours
Regular working hours

Tech stack

Agile Methodology Amazon Web Services Amazon S3 Microsoft Azure Big Data Information Engineering Extract Transform Load (ETL) Data Warehousing Python (Programming Language) Performance Tuning Scrum Methodology Azure Data Lake
+7 more
SQL Databases Azure Data Factory Data Lakes Pyspark Information Technology Data Pipelines Databricks

Job description

We are seeking a skilled Data Engineer to join our team. The successful candidate will be responsible for development and optimization of data pipelines, implementing robust data checks, and ensuring the accuracy and integrity of data flows. This role is critical in supporting data-driven decision-making processes, especially in the context of our insurance-focused business operations, * Collaborate with data analysts, reporting team and business advisors to gather requirements and define data models that effectively support business requirements

  • Develop and maintain scalable and efficient data pipelines to ensure seamless data flow across various systems adddress any issues or bottlenecks in existing pipelines.
  • Implement robust data checks to ensure the accuracy and integrity of data. Summarize and validate large datasets to ensure they meet quality standards.
  • Monitor data jobs for successful completion. Troubleshoot and resolve any issues that arise to minimize downtime and ensure continuity of data processes.
  • Regularly review and audit data processes and pipelines to ensure compliance with internal standards and regulatory requirements
  • Familiar with working on Agile methodologies - scrum, sprint planning, backlog refinement etc.

Requirements

  • 7-12 years experience on Data Engineering role working with Databricks and any Cloud technologies- AWS , Azure , Databricks.
  • Bachelor’s degree in computer science, Information Technology, or related field.
  • Strong proficiency in PySpark, Python, SQL.
  • Strong experience in data modeling, ETL/ELT pipeline development, and automation
  • Hands-on experience with performance tuning of data pipelines and workflows
  • Proficient in working on any cloud components Azure Data Factory, Azure DataBricks, Azure Data Lake , S3, Lamda.
  • Experience with data modeling, ETL processes, Delta Lake and data warehousing.
  • Experience on Delta Live Tables, Autoloader & Unity Catalog.
  • Preferred - Knowledge of the insurance industry and its data requirements.
  • Strong analytical skills with the ability to collect, organize, analyze, and disseminate significant amounts of information with attention to detail and accuracy.
  • Excellent communication and problem-solving skills to work effectively with diverse teams
  • Excellent problem-solving skills and ability to work under tight deadlines.

Benefits & conditions

Posted Yesterday Remote or Hybrid Hiring Remotely in United States 110K-130K Annually Senior level Remote or Hybrid Hiring Remotely in United States 110K-130K Annually Senior level Design, build, and optimize scalable ETL/ELT data pipelines using Databricks and cloud platforms. Implement data quality checks, monitor and troubleshoot jobs, audit pipelines for compliance, and collaborate with analysts and business stakeholders to define data models and support reporting. The summary above was generated by AI

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on fa-ewjt-saasfaprod1.fa.ocs.oraclecloud.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all