> Markdown version of [/jobs/ext/1346445-data-and-cloud-engineer](https://www.wearedevelopers.com/jobs/ext/1346445-data-and-cloud-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data and Cloud Engineer - **Company:** EXL SERVICE - **Location:** New York, NY, United States (Remote available) - **Experience:** Experienced - **Salary:** $190,000.0 - $193,000.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Airflow, Amazon Web Services, Data Analysis, Big Data, Cloud Computing, Cloud Engineering, Cloudera Impala, Databases, Continuous Integration, Information Engineering, Data Infrastructure, Extract Transform Load (ETL), Data Warehousing, DevOps, Distributed Computing Environment, Django Web Framework, Github, Apache Hadoop, Hadoop Distributed File System, Apache Hive, Python (Programming Language), Object-Oriented Software Development, Windows PowerShell, Scrum Methodology, Software Engineering, SQL Databases, Sqoop, Strategies of Testing, Apache Zookeeper, Google Cloud, Cloud Platform System, Apache Yarn, Flask (Web Framework), Snowflake, Apache Spark, Data Lakes, Information Technology, Apache Kafka, Data Management, Machine Learning Operations, Data Pipelines, Databricks - **Published:** July 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=88e13cb45435ccdd ## About the Role Qualifications: Requires a Bachelor's degree Computer Science, Engineering, or a directly related field plus five (5) years of software engineering experience. Experience must include: Two (2) years of experience in: SQL scripting; Big data technologies: HDFS, YARN, Spark, Hive, Sqoop, Impala, Airflow, Zookeeper, and Kafka; Enterprise GitHub: branch, release, DevOps, and CI/CD pipeline; and Data engineering or ETL development. Three (3) months of experience in: Snowflake or other cloud-based data platforms (AWS, GCP, Databricks) and Python web frameworks: Django or Flask. Employer will also accept a Master's degree plus three (3) years of engineering experience in lieu of Bachelor's plus five (5) years of software engineering experience. ## Description Job Description: Design, implement and deliver large scale enterprise applications using big data open-source solutions such as Apache Hadoop, Apache Spark, Kafka, and Elastic Search. Implement data pipelines and data driven applications using Python on distributed computing frameworks like Hadoop, Apache Spark, etc. * Responsibilities: Use AWS services and GCP services to build data pipelines and migrate on-prem data pipelines and data applications to Cloud infrastructure. * Proficient in using AWS cloud-based services to implement batch and online steaming applications. * Work closely with Data science teams to integrate data, algorithms into data lake systems and automate different Machine Learning workflows and assist with data infrastructure needs. * Experience in working both, On-Prem and Cloud. * Design and implement efficient data pipelines (ETLs) in order to integrate data from a variety of sources into Data Warehouse. * Design and implement data model changes that align with warehouse standards. * Design and implement backfill or other warehouse data management processes. * Develop and execute testing strategies to ensure high quality warehouse data. * Provide documentation, training, and consulting for data warehouse users. * Perform requirement and data analysis in order to support warehouse project definition. * Excellent database troubleshooting skills. * Working technical knowledge of PowerShell. * Strong object-oriented design and analysis skills. * Verbal and written communication skills and the ability to interact professionally with a diverse group, executives, managers, and subject matter experts. * Working in agile methodology, involve in Grooming, Sprint Planning, and daily Scrum meetings. * Position may work at various and unanticipated worksites throughout the United States. Telecommuting permitted. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)