Lead Data Engineer
Strategic Staffing Solutions
Chandler, AZ, United States
about 2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source
Tech stack
Agile Methodology
Artificial Intelligence
Airflow
Amazon Web Services
Amazon S3
CA Workload Automation Ae
Microsoft Azure
Big Data
BigQuery
Code Coverage
Code Review
Information Engineering
+31 more
Data Integration
Extract Transform Load (ETL)
Data Warehousing
Relational Databases
Distributed Systems
Apache Hadoop
Apache Hive
Python (Programming Language)
Machine Learning
Oracle Databases
Scrum Methodology
Shell Script
Simple Data Format
PL-SQL
SQL Databases
Teradata SQL
Unix Commands
Data Processing
Scripting
Freeform SQL
Google Cloud
Large Language Models
Apache Spark
Parallel Computation
Ab Initio
Backend
Apache Kafka
Data Management
Virtual Agents
Stream Processing
Data Pipelines
Job description
- Role will be working with the enterprise information warehouse. They are looking for a specialized ETL backend developer to help with additional work they’ve taken on related to the migration to GCP
- Part of the Data & Insights team, they build data platforms to deliver strategic information and insights to the business
- The role will be analyzing new development requests, designing and developing new ETL processes, testing, and putting them into production.
In this role, you will:
- Lead moderately complex initiatives within Technology and contribute to large scale data processing framework initiatives related to enterprise strategy deliverables
- Build and maintain optimized and highly available data pipelines that facilitate deeper analysis and reporting
- Review and analyze moderately complex business, operational or technical challenges that require an in-depth evaluation of variable factors
- Oversee the data integration work, including developing a data model, maintaining a data warehouse and analytics environment, and writing scripts for data integration and analysis
- Resolve moderately complex issues and lead teams to meet data engineering deliverables while leveraging solid understanding of data information policies, procedures and compliance requirements
- Collaborate and consult with colleagues and managers to resolve data engineering issues and achieve strategic goals.
Requirements
Do you have experience in Teradata?, * 6+ years of Data Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
- 6+ years of strong experience with Ab Initio, Spark Core, Spark SQL, Data Frames, Datasets
- 3+ Experience on Python or other scripting languages
- Hands-on experience with GCP or other cloud platforms (AWS/Azure)
- 3+ years of data warehouse experience, SQL experience (Teradata, SQL, PL/SQL). Understand and write the complex SQL queries and have in-depth understanding of oracle database
- Strong experience with Teradata, Hadoop, Hive, Autosys, Airflow, BigQuery SparkFlow
- Good understanding of distributed system and parallel processing, spark architecture in depth
- Ability to conduct code reviews with focus on testability and code coverage.
- Should have executed end to end projects in data engineering
- Should have experience working in an Agile environment using Scrum.
- Fluent in communication
Desired Qualifications:
- Prior experience of leading the teams and running the major programs is needed.
- Good knowledge of non-relational DBMS platforms, (MongoDBj), real-time data streaming via Kafka
- Working knowledge in developing end to end ML/AI pipeline in Apache Spark, Sparkling Waters, H2O, experience in working on LLMs. Exposer to Agentic AI frameworks.
- Ability to work with Data Engineers in discovering and optimizing bottlenecks in the AI/ML pipeline for real-time or near-real-time applications that consumes large throughput of data
- Perform various complex activities related to statistical/machine learning. Provide analytical support for developing, evaluating, implementing, monitoring and executing models across business verticals using emerging technologies including but not limited to Python, Agentic AI frameworks like ADK, GCP vertex AI and NLP.
- Proficient in large scala data processing, Hands-on experience with different file formats, very good understanding of Unix command & Shell scripting.
- Familiarity with CI/CD pipelines, infrastructure as code and cloud deployment best practices
“Beware of scams. S3 never asks for money during its onboarding process.”
Application Question(s):
- Are you on OPT/CPT visa? If yes, pls withdraw your application as we cannot present this visa type due to client policy.
- Are you able to work on our W2? this is not a c2c or 1099 work type
Experience:
- Ab Initio: 5 years (Required)
- Google Cloud Platform: 5 years (Required)
- Spark: 5 years (Required)
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
ER
Erin Rifkin
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
about 1 year ago
LM
Luis Minvielle
How to Become an AI Engineer
over 2 years ago
BB
Benedikt Bischof
Making Data Warehouses Fast: A Developer’s Story
about 4 years ago
DS
Dhannush Subramani
Top Big Data Technologies That You Need to Know
about 4 years ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago