Lead Data Engineer

Strategic Staffing Solutions
Chandler, United States of America
24 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior

Job location

Chandler, United States of America

Tech stack

Agile Methodologies
Artificial Intelligence
Airflow
Amazon Web Services (AWS)
Amazon Web Services (AWS)
CA Workload Automation Ae
Azure
Big Data
Google BigQuery
Code Coverage
Code Review
Information Engineering
Data Integration
ETL
Data Warehousing
Relational Databases
Distributed Systems
Hadoop
Hive
Python
Machine Learning
Oracle
Scrum
Shell Script
Simple Data Format
PL-SQL
SQL Databases
Teradata
Unix Commands
Data Processing
Scripting (Bash/Python/Go/Ruby)
Freeform SQL
Google Cloud Platform
Large Language Models
Spark
Parallel Computation
Ab Initio
Backend
Kafka
Data Management
Virtual Agents
Stream Processing
Data Pipelines

Job description

  • Role will be working with the enterprise information warehouse. They are looking for a specialized ETL backend developer to help with additional work they've taken on related to the migration to GCP
  • Part of the Data & Insights team, they build data platforms to deliver strategic information and insights to the business
  • The role will be analyzing new development requests, designing and developing new ETL processes, testing, and putting them into production.

In this role, you will:

  • Lead moderately complex initiatives within Technology and contribute to large scale data processing framework initiatives related to enterprise strategy deliverables
  • Build and maintain optimized and highly available data pipelines that facilitate deeper analysis and reporting
  • Review and analyze moderately complex business, operational or technical challenges that require an in-depth evaluation of variable factors
  • Oversee the data integration work, including developing a data model, maintaining a data warehouse and analytics environment, and writing scripts for data integration and analysis
  • Resolve moderately complex issues and lead teams to meet data engineering deliverables while leveraging solid understanding of data information policies, procedures and compliance requirements
  • Collaborate and consult with colleagues and managers to resolve data engineering issues and achieve strategic goals.

Requirements

Do you have experience in Teradata?, * 6+ years of Data Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education

  • 6+ years of strong experience with Ab Initio, Spark Core, Spark SQL, Data Frames, Datasets
  • 3+ Experience on Python or other scripting languages
  • Hands-on experience with GCP or other cloud platforms (AWS/Azure)
  • 3+ years of data warehouse experience, SQL experience (Teradata, SQL, PL/SQL). Understand and write the complex SQL queries and have in-depth understanding of oracle database
  • Strong experience with Teradata, Hadoop, Hive, Autosys, Airflow, BigQuery SparkFlow
  • Good understanding of distributed system and parallel processing, spark architecture in depth
  • Ability to conduct code reviews with focus on testability and code coverage.
  • Should have executed end to end projects in data engineering
  • Should have experience working in an Agile environment using Scrum.
  • Fluent in communication

Desired Qualifications:

  • Prior experience of leading the teams and running the major programs is needed.
  • Good knowledge of non-relational DBMS platforms, (MongoDBj), real-time data streaming via Kafka
  • Working knowledge in developing end to end ML/AI pipeline in Apache Spark, Sparkling Waters, H2O, experience in working on LLMs. Exposer to Agentic AI frameworks.
  • Ability to work with Data Engineers in discovering and optimizing bottlenecks in the AI/ML pipeline for real-time or near-real-time applications that consumes large throughput of data
  • Perform various complex activities related to statistical/machine learning. Provide analytical support for developing, evaluating, implementing, monitoring and executing models across business verticals using emerging technologies including but not limited to Python, Agentic AI frameworks like ADK, GCP vertex AI and NLP.
  • Proficient in large scala data processing, Hands-on experience with different file formats, very good understanding of Unix command & Shell scripting.
  • Familiarity with CI/CD pipelines, infrastructure as code and cloud deployment best practices

"Beware of scams. S3 never asks for money during its onboarding process."

Application Question(s):

  • Are you on OPT/CPT visa? If yes, pls withdraw your application as we cannot present this visa type due to client policy.
  • Are you able to work on our W2? this is not a c2c or 1099 work type

Experience:

  • Ab Initio: 5 years (Required)
  • Google Cloud Platform: 5 years (Required)
  • Spark: 5 years (Required)

Apply for this position