> Markdown version of [/jobs/ext/76302-data-integration-engineer](https://www.wearedevelopers.com/jobs/ext/76302-data-integration-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Integration Engineer - **Company:** Tata Consultancy Services Limited - **Location:** Sunnyvale, CA, United States - **Salary:** $80,000.0 - $140,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Airflow, Batch Processing, Computer Programming, Data Validation, Information Engineering, Data Integration, Extract Transform Load (ETL), Data Sharing, Data Visualization, Data Warehousing, Cursor (Graphical User Interface Elements), Database Queries, Python (Programming Language), NumPy, Performance Tuning, Tensorflow, Search Technologies, SQL Stored Procedures, SQL Databases, Tableau (Software), Text Mining, Unstructured Data, Data Ingestion, Pytorch, Large Language Models, Snowflake, Pandas, Scikit Learn, Apache Kafka, Text Analysis, Data Pipelines, Sql Tuning - **Published:** May 16, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=df16bce6ff0b45b9 ## About the Role Do you have experience in Text mining?, * Snowflake Data Engineering - o Design and implement enterprise-grade data pipelines using Snowflake, including ingestion and transformation o Must be strong in both Core and Semantic aspects o Develop complex SQL transformations, stored procedures and Dynamic tables inside Snowflake to enable near real-time and batch processing o Implement Snowflake data sharing, data marketplace integrations o Engineer Snowpipe and Kafka-to-Snowflake streaming ingestion pipelines also handling high throughput event data at scale o Optimize Snowflake cluster performance - virtual warehouse sizing, query profiling, clustering keys o Architecture, design aspects, performance tuning, time travel, warehouse concepts - scaling, clustering, micro-partitioning o Experience with SnowSQL, Snowpipe * Data Integration aspects - o Design and maintain end-to-end ETL/ELT pipelines using Apache Airflow o Experience in building reusable parameterized data ingestion pipelines/frameworks is beneficial. o Thorough on data quality checks * AI and Data Science - o Integrate AI/LLMs with data pipelines via Python UDFs or API callouts - enabling text analytics, semantic search and GEN-AI augmented workflows o Experience with Python based frameworks - scikit learn, PyTorch, TensorFlow o Experience with NLP and text-mining techniques on unstructured data to identify actionable information o Time-series forecasting, anomaly detection and propensity modeling * Experience with Data Visualization aspects * Hands-on experience with writing Complex queries using - Joins, Self Joins, Views, Materialized Views, Cursor also Recursive, use of GROUP BY, PARTITION BY functions / SQL Performance tuning * Hands-on experience with ETL and Dimensional Data Modelling - Slowly Changing Dimensions (SCD - Type 1, 2, 3) o Good understanding of concepts like schema types, table types - fact-dimension etc. like how to design a dimension vs fact, design considerations factored etc. * Proficiency in Python scripting/programming - using Pandas, PyParsing, Airflow. o Pandas, Tableau server modules, Numpy, Datetime, Apache Airflow related modules, APIs o Data Pipeline automation o Strong Python programming skills * Actively participating in discussions with business to understand requirements, perform thorough impact analysis and provide suitable solutions. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers)