Lead Data Engineer

Strategic Staffing Solutions
Chandler, AZ, United States
about 2 months ago
Apply on indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Airflow Amazon Web Services Amazon S3 CA Workload Automation Ae Microsoft Azure Big Data BigQuery Code Coverage Code Review Information Engineering
+31 more
Data Integration Extract Transform Load (ETL) Data Warehousing Relational Databases Distributed Systems Apache Hadoop Apache Hive Python (Programming Language) Machine Learning Oracle Databases Scrum Methodology Shell Script Simple Data Format PL-SQL SQL Databases Teradata SQL Unix Commands Data Processing Scripting Freeform SQL Google Cloud Large Language Models Apache Spark Parallel Computation Ab Initio Backend Apache Kafka Data Management Virtual Agents Stream Processing Data Pipelines

Job description

  • Role will be working with the enterprise information warehouse. They are looking for a specialized ETL backend developer to help with additional work they’ve taken on related to the migration to GCP
  • Part of the Data & Insights team, they build data platforms to deliver strategic information and insights to the business
  • The role will be analyzing new development requests, designing and developing new ETL processes, testing, and putting them into production.

In this role, you will:

  • Lead moderately complex initiatives within Technology and contribute to large scale data processing framework initiatives related to enterprise strategy deliverables
  • Build and maintain optimized and highly available data pipelines that facilitate deeper analysis and reporting
  • Review and analyze moderately complex business, operational or technical challenges that require an in-depth evaluation of variable factors
  • Oversee the data integration work, including developing a data model, maintaining a data warehouse and analytics environment, and writing scripts for data integration and analysis
  • Resolve moderately complex issues and lead teams to meet data engineering deliverables while leveraging solid understanding of data information policies, procedures and compliance requirements
  • Collaborate and consult with colleagues and managers to resolve data engineering issues and achieve strategic goals.

Requirements

Do you have experience in Teradata?, * 6+ years of Data Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education

  • 6+ years of strong experience with Ab Initio, Spark Core, Spark SQL, Data Frames, Datasets
  • 3+ Experience on Python or other scripting languages
  • Hands-on experience with GCP or other cloud platforms (AWS/Azure)
  • 3+ years of data warehouse experience, SQL experience (Teradata, SQL, PL/SQL). Understand and write the complex SQL queries and have in-depth understanding of oracle database
  • Strong experience with Teradata, Hadoop, Hive, Autosys, Airflow, BigQuery SparkFlow
  • Good understanding of distributed system and parallel processing, spark architecture in depth
  • Ability to conduct code reviews with focus on testability and code coverage.
  • Should have executed end to end projects in data engineering
  • Should have experience working in an Agile environment using Scrum.
  • Fluent in communication

Desired Qualifications:

  • Prior experience of leading the teams and running the major programs is needed.
  • Good knowledge of non-relational DBMS platforms, (MongoDBj), real-time data streaming via Kafka
  • Working knowledge in developing end to end ML/AI pipeline in Apache Spark, Sparkling Waters, H2O, experience in working on LLMs. Exposer to Agentic AI frameworks.
  • Ability to work with Data Engineers in discovering and optimizing bottlenecks in the AI/ML pipeline for real-time or near-real-time applications that consumes large throughput of data
  • Perform various complex activities related to statistical/machine learning. Provide analytical support for developing, evaluating, implementing, monitoring and executing models across business verticals using emerging technologies including but not limited to Python, Agentic AI frameworks like ADK, GCP vertex AI and NLP.
  • Proficient in large scala data processing, Hands-on experience with different file formats, very good understanding of Unix command & Shell scripting.
  • Familiarity with CI/CD pipelines, infrastructure as code and cloud deployment best practices

“Beware of scams. S3 never asks for money during its onboarding process.”

Application Question(s):

  • Are you on OPT/CPT visa? If yes, pls withdraw your application as we cannot present this visa type due to client policy.
  • Are you able to work on our W2? this is not a c2c or 1099 work type

Experience:

  • Ab Initio: 5 years (Required)
  • Google Cloud Platform: 5 years (Required)
  • Spark: 5 years (Required)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

1:12 min

Choosing TypeScript for complex backend applications

Maximilian Otto Maximilian Otto · World Congress 2024

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all