AI Data Engineer

Diverse Lynx LLC
Indianapolis, IN, United States
5 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$101,500.0 - $169,100.0
Working hours
Regular working hours

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services Business Analytics Applications Data Analysis Microsoft Azure Big Data Cloud Database Continuous Integration Information Engineering Data Infrastructure Extract Transform Load (ETL)
+34 more
Data Transformation Data Warehousing Relational Databases Database Queries DevOps Github Python (Programming Language) PostgreSQL Lynx NoSQL Power BI Cloud Services Azure Data Lake SQL Stored Procedures SQL Databases Data Streaming Tableau (Software) Unstructured Data Scripting Data Ingestion Sql Optimization Snowflake Microsoft Fabric Pyspark Information Technology Deployment Automation Star Schema SAP S/4HANA Azure Synapse Analytics Software Version Control Data Pipelines Serverless Computing Amazon Redshift Databricks

Job description

Responsible for developing, deploying and optimizing scalable data pipelines, data models, and analytics solutions to support enterprise-wide business insights and decision-making. This role requires expertise in cloud data engineering, data warehousing, advanced SQL and Python/PySpark transformations , with the ability to translate data into actionable business recommendations. The ideal candidate brings deep technical experience, strong analytical skills, and cross-functional collaboration abilities-preferably within the pharmaceutical or healthcare domain. Core Responsibilities Data Engineering & Pipeline Development · Develop and maintain scalable ETL/ELT pipelines using Python, PySpark, and SQL for structured and unstructured data ingestion · Build robust data infrastructure on cloud platforms (Azure Fabric, Synapse, Databricks, AWS) · Implement end-to-end automation for data ingestion, streaming, scheduling, and monitoring across Azure Functions, Data Factory, ADLS Gen2, and Power BI · Optimize large-scale big data architectures for performance and scalability Data Modelling & Architecture · Design and maintain multidimensional models (Star/Snowflake Schema) including fact and dimension tables, views, and stored procedures · Manage complex datasets ensuring alignment with functional and non-functional requirements · Learn foundational SAP S4/HANA and BW4 data models to support financial analytics DevOps & Deployment · Utilize GitHub and Azure DevOps for CI/CD, version control, and automated deployments · Manage artifact deployment across environments with risk assessments and impact analysis · Implement process improvements and workflow automation Analytics & Business Partnership · Partner with business stakeholders to gather requirements and translate them into analytical solutions · Develop analytics and reporting solutions supporting Global Finance transformation strategy · Develop dashboards and BI tools using Power BI/Tableau, translating outcomes into actionable insights with KPIs · Troubleshoot data-related issues and perform root cause analysis

Requirements

· 5-8 years of hands-on experience in data engineering and cloud data warehousing (Azure Synapse, Fabric, Databricks, Redshift, or Snowflake) · Expert SQL proficiency across relational databases and cloud data warehouses, plus strong Python and PySpark skills for data transformation and pipeline development · Data modelling expertise including dimensional modelling (facts, dims), normalization/de-normalization, views, and stored procedures for automation · End-to-end ownership of project delivery with proven ability to translate business problems into scalable analytical solutions · Big data architecture experience building and optimizing large-scale pipelines across structured and unstructured datasets · Visualization expertise with Power BI and/or Tableau to present complex findings clearly · Strong documentation and communication skills to support operational standards and collaborate with cross-functional teams · Experience with root cause analysis, continuous improvement, and Agile delivery methodologies Preferred Qualifications · Experience with SQL and NoSQL databases (AWS Redshift, Postgres, Databricks etc.) · Pharmaceutical or healthcare dataset experience · Cloud expertise in AWS/Azure; Microsoft Fabric experience is a strong advantage · CI/CD experience using GitHub or Azure DevOps · Proficiency in Python/PySpark, R, Scala, or additional scripting languages Education · Bachelor’s or Master’s degree in Technology, Computer Science, Engineering, or related field Competencies · Strong analytical and problem-solving skills · Excellent communication and stakeholder management · High attention to detail and quality · Ability to thrive in fast-paced, dynamic environments · Strategic and critical thinking · Cross-functional collaboration and alignment Diverse Lynx LLC is an Equal Employment Opportunity employer. All qualified applicants will receive due consideration for employment without any discrimination. All applicants will be evaluated solely on the basis of their ability, competence and their proven capability to perform the functions outlined in the corresponding role. We promote and support a diverse workforce across all levels in the company.

About the company

Cox Automotive

  • Indianapolis, IN
  • $101,500-169,100 per year Cox Automotive is deploying enterprise AI capabilities on AWS Quick across the enterprise, helping teams work more effectively in their day-to-day operations. We are looking for …

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:57 min

Introduction to the Lynx cross-platform UI framework

Xuan Huang Xuan Huang · World Congress 2025

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all