> Markdown version of [/jobs/ext/2632485-sr-data-engineer-intl-india](https://www.wearedevelopers.com/jobs/ext/2632485-sr-data-engineer-intl-india). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Data Engineer - INTL India - **Company:** Insight Global - **Location:** Pittsburgh, PA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Third Normal Form, Artificial Intelligence, Amazon Web Services, Microsoft Azure, Cloud Engineering, Continuous Integration, Data Architecture, Information Engineering, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Relational Databases, DevOps, Github, Python (Programming Language), Metadata, Meta-Data Management, MongoDB, NoSQL, DataOps, Scala (Programming Language), SQL Databases, Data Streaming, Talend, Enterprise Data Management, Google Cloud, Data Ingestion, Snowflake, Multi-Agent Systems, Apache Spark, Event Driven Architecture, Containerization, Kubernetes, Infrastructure Automation Frameworks, Data Lineage, Graphql, Data Management, Machine Learning Operations, Data Lakehouse, Restful APIs, Software Version Control, Docker, Databricks, Web Api - **Published:** August 14, 2026 - **Apply:** https://dejobs.org/x/x/25E6918635B04FC1941D03CF30C43811/job/ ## About the Role 5-8+ years of hands-on experience in data architecture and engineering delivery. Proven success in building modern data platforms on cloud (Azure, AWS, GCP). Deep knowledge of data lakehouse architectures (e.g., Databricks, Fabric). Proficiency with Python, SQL, Spark, and orchestration frameworks. Experience with ETL/ELT tools (e.g., Informatica, Talend, Fivetran) and containerization (Docker, Kubernetes). Strong background in Data Modeling (ERD, star/snowflake, canonical models). Familiarity with REST APIs, GraphQL, and event-driven design. Demonstrated experience integrating AI/ML and GenAI components into data platforms. Exposure to DataOps and DevOps practices for CI/CD and platform automation. * Working knowledge of code management and CI/CD systems (Azure DevOps or GitHub). * Familiarity with NoSQL databases (e.g., MongoDB). * Exposure to IoT Data Standards like Project Haystack, Brick Schema, Real Estate Core. ## Description Define and continuously evolve the data architecture across the Digital Products and Platforms Translate business and technical goals into scalable and resilient platform designs. Own and maintain architectural roadmaps, standards, and decision frameworks. Act as the bridge between architects, SME/Analysts, data engineers, and analytics teams to ensure alignment and compliance with platform standards. Data Engineering & Platform Delivery Design and implement modern ELT/ETL pipelines using tools like Spark, Python, SQL, Scala, and cloud-native components (e.g., Databricks). Design AI ready Data Models (e.g., RAG, agent orchestration, multimodal pipelines) with working reference implementations. Build reusable AI components, templates, and accelerators to enable consistent adoption across teams. Implement and optimize scalable, secure, and resilient AI pipelines, aligned with enterprise data and governance standards. Lead PoC-to-production transitions, ensuring operational readiness, observability, and cost controls Demonstrated success designing and deploying RAG architectures, including vector stores, embedding strategies, chunking logic, semantic retrieval, and hybrid search Deep Hands-on experience with at least one of the following Hyperscaler AI / Data Platforms: Azure, AWS Design and implement Data Modelling using Relational DB/NoSQL DB's. Proven Hand-on experience in Databricks & Unity Catalog. Manage data ingestion from heterogeneous sources including ERP, CRM, IoT, and third-party APIs. Guide hands-on development of robust, reusable, and automated data flows. Governance, Metadata, and Quality Implement and enforce data governance frameworks including data lineage, metadata management, and access controls. Develop data models (ERDs, dimensional and 3NF) and define canonical data representations. Collaboration & Leadership Review solution designs and provide architectural guidance to engineering teams. Mentor technical staff while fostering best practices and continuous improvement. Collaborate with DevOps to embed CI/CD, version control, and environment automation across the data lifecycle. Continuously assess and improve platform reliability, scalability, performance, and cost-efficienc We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts)