> Markdown version of [/jobs/ext/2714407-ai-data-engineer](https://www.wearedevelopers.com/jobs/ext/2714407-ai-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Data Engineer - **Company:** Diverse Lynx LLC - **Location:** Indianapolis, IN, United States - **Experience:** Expert - **Salary:** $101,500.0 - $169,100.0 - **Contract:** Temporary to permanent - **Skills:** Agile Methodology, Artificial Intelligence, Amazon Web Services, Business Analytics Applications, Data Analysis, Microsoft Azure, Big Data, Cloud Database, Continuous Integration, Information Engineering, Data Infrastructure, Extract Transform Load (ETL), Data Transformation, Data Warehousing, Relational Databases, Database Queries, DevOps, Github, Python (Programming Language), PostgreSQL, Lynx, NoSQL, Power BI, Cloud Services, Azure Data Lake, SQL Stored Procedures, SQL Databases, Data Streaming, Tableau (Software), Unstructured Data, Scripting, Data Ingestion, Sql Optimization, Snowflake, Microsoft Fabric, Pyspark, Information Technology, Deployment Automation, Star Schema, SAP S/4HANA, Azure Synapse Analytics, Software Version Control, Data Pipelines, Serverless Computing, Amazon Redshift, Databricks - **Published:** September 4, 2026 - **Apply:** https://www.careerjet.com/job/use824897f7ff605be01f5e509738df114/eaa ## About the Role · 5-8 years of hands-on experience in data engineering and cloud data warehousing (Azure Synapse, Fabric, Databricks, Redshift, or Snowflake) · Expert SQL proficiency across relational databases and cloud data warehouses, plus strong Python and PySpark skills for data transformation and pipeline development · Data modelling expertise including dimensional modelling (facts, dims), normalization/de-normalization, views, and stored procedures for automation · End-to-end ownership of project delivery with proven ability to translate business problems into scalable analytical solutions · Big data architecture experience building and optimizing large-scale pipelines across structured and unstructured datasets · Visualization expertise with Power BI and/or Tableau to present complex findings clearly · Strong documentation and communication skills to support operational standards and collaborate with cross-functional teams · Experience with root cause analysis, continuous improvement, and Agile delivery methodologies Preferred Qualifications · Experience with SQL and NoSQL databases (AWS Redshift, Postgres, Databricks etc.) · Pharmaceutical or healthcare dataset experience · Cloud expertise in AWS/Azure; Microsoft Fabric experience is a strong advantage · CI/CD experience using GitHub or Azure DevOps · Proficiency in Python/PySpark, R, Scala, or additional scripting languages Education · Bachelor's or Master's degree in Technology, Computer Science, Engineering, or related field Competencies · Strong analytical and problem-solving skills · Excellent communication and stakeholder management · High attention to detail and quality · Ability to thrive in fast-paced, dynamic environments · Strategic and critical thinking · Cross-functional collaboration and alignment Diverse Lynx LLC is an Equal Employment Opportunity employer. All qualified applicants will receive due consideration for employment without any discrimination. All applicants will be evaluated solely on the basis of their ability, competence and their proven capability to perform the functions outlined in the corresponding role. We promote and support a diverse workforce across all levels in the company. ## Description Responsible for developing, deploying and optimizing scalable data pipelines, data models, and analytics solutions to support enterprise-wide business insights and decision-making. This role requires expertise in cloud data engineering, data warehousing, advanced SQL and Python/PySpark transformations , with the ability to translate data into actionable business recommendations. The ideal candidate brings deep technical experience, strong analytical skills, and cross-functional collaboration abilities-preferably within the pharmaceutical or healthcare domain. Core Responsibilities Data Engineering & Pipeline Development · Develop and maintain scalable ETL/ELT pipelines using Python, PySpark, and SQL for structured and unstructured data ingestion · Build robust data infrastructure on cloud platforms (Azure Fabric, Synapse, Databricks, AWS) · Implement end-to-end automation for data ingestion, streaming, scheduling, and monitoring across Azure Functions, Data Factory, ADLS Gen2, and Power BI · Optimize large-scale big data architectures for performance and scalability Data Modelling & Architecture · Design and maintain multidimensional models (Star/Snowflake Schema) including fact and dimension tables, views, and stored procedures · Manage complex datasets ensuring alignment with functional and non-functional requirements · Learn foundational SAP S4/HANA and BW4 data models to support financial analytics DevOps & Deployment · Utilize GitHub and Azure DevOps for CI/CD, version control, and automated deployments · Manage artifact deployment across environments with risk assessments and impact analysis · Implement process improvements and workflow automation Analytics & Business Partnership · Partner with business stakeholders to gather requirements and translate them into analytical solutions · Develop analytics and reporting solutions supporting Global Finance transformation strategy · Develop dashboards and BI tools using Power BI/Tableau, translating outcomes into actionable insights with KPIs · Troubleshoot data-related issues and perform root cause analysis ## Related Videos - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Lynx: Native for More](https://www.wearedevelopers.com/videos/1445-lynx-native-for-more) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)