> Markdown version of [/jobs/ext/1111793-lead-data-engineer](https://www.wearedevelopers.com/jobs/ext/1111793-lead-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data Engineer - **Company:** Strategic Staffing Solutions - **Location:** Chandler, AZ, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, CA Workload Automation Ae, Microsoft Azure, Big Data, BigQuery, Code Coverage, Code Review, Information Engineering, Data Integration, Extract Transform Load (ETL), Data Warehousing, Relational Databases, Distributed Systems, Apache Hadoop, Apache Hive, Python (Programming Language), Machine Learning, Oracle Databases, Scrum Methodology, Shell Script, Simple Data Format, PL-SQL, SQL Databases, Teradata SQL, Unix Commands, Data Processing, Scripting, Freeform SQL, Google Cloud, Large Language Models, Apache Spark, Parallel Computation, Ab Initio, Backend, Apache Kafka, Data Management, Virtual Agents, Stream Processing, Data Pipelines - **Published:** June 30, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=93faca02cef5d21a ## About the Role Do you have experience in Teradata?, * 6+ years of Data Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education * 6+ years of strong experience with Ab Initio, Spark Core, Spark SQL, Data Frames, Datasets * 3+ Experience on Python or other scripting languages * Hands-on experience with GCP or other cloud platforms (AWS/Azure) * 3+ years of data warehouse experience, SQL experience (Teradata, SQL, PL/SQL). Understand and write the complex SQL queries and have in-depth understanding of oracle database * Strong experience with Teradata, Hadoop, Hive, Autosys, Airflow, BigQuery SparkFlow * Good understanding of distributed system and parallel processing, spark architecture in depth * Ability to conduct code reviews with focus on testability and code coverage. * Should have executed end to end projects in data engineering * Should have experience working in an Agile environment using Scrum. * Fluent in communication Desired Qualifications: * Prior experience of leading the teams and running the major programs is needed. * Good knowledge of non-relational DBMS platforms, (MongoDBj), real-time data streaming via Kafka * Working knowledge in developing end to end ML/AI pipeline in Apache Spark, Sparkling Waters, H2O, experience in working on LLMs. Exposer to Agentic AI frameworks. * Ability to work with Data Engineers in discovering and optimizing bottlenecks in the AI/ML pipeline for real-time or near-real-time applications that consumes large throughput of data * Perform various complex activities related to statistical/machine learning. Provide analytical support for developing, evaluating, implementing, monitoring and executing models across business verticals using emerging technologies including but not limited to Python, Agentic AI frameworks like ADK, GCP vertex AI and NLP. * Proficient in large scala data processing, Hands-on experience with different file formats, very good understanding of Unix command & Shell scripting. * Familiarity with CI/CD pipelines, infrastructure as code and cloud deployment best practices "Beware of scams. S3 never asks for money during its onboarding process." Application Question(s): * Are you on OPT/CPT visa? If yes, pls withdraw your application as we cannot present this visa type due to client policy. * Are you able to work on our W2? this is not a c2c or 1099 work type Experience: * Ab Initio: 5 years (Required) * Google Cloud Platform: 5 years (Required) * Spark: 5 years (Required) ## Description * Role will be working with the enterprise information warehouse. They are looking for a specialized ETL backend developer to help with additional work they've taken on related to the migration to GCP * Part of the Data & Insights team, they build data platforms to deliver strategic information and insights to the business * The role will be analyzing new development requests, designing and developing new ETL processes, testing, and putting them into production. In this role, you will: * Lead moderately complex initiatives within Technology and contribute to large scale data processing framework initiatives related to enterprise strategy deliverables * Build and maintain optimized and highly available data pipelines that facilitate deeper analysis and reporting * Review and analyze moderately complex business, operational or technical challenges that require an in-depth evaluation of variable factors * Oversee the data integration work, including developing a data model, maintaining a data warehouse and analytics environment, and writing scripts for data integration and analysis * Resolve moderately complex issues and lead teams to meet data engineering deliverables while leveraging solid understanding of data information policies, procedures and compliance requirements * Collaborate and consult with colleagues and managers to resolve data engineering issues and achieve strategic goals. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)