Data Engineer

Everforth Apex
Santa Ana, United States of America
yesterday

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior

Job location

Remote
Santa Ana, United States of America

Tech stack

Artificial Intelligence
Airflow
Data analysis
Apache HTTP Server
Google BigQuery
Cloud Database
Cloud Engineering
Cloud Storage
Code Review
Computer Programming
Continuous Delivery
Continuous Integration
Directed Acyclic Graph (Directed Graphs)
Data Architecture
Information Engineering
Data Governance
Data Infrastructure
Data Integration
ETL
Data Profiling
Data Visualization
Relational Databases
Cursor (Graphical User Interface Elements)
Database Queries
Programming Tools
Data Flow Control
Github
Hive
Python
Operational Databases
Performance Tuning
Query Optimization
Power BI
Cloudera
Software Deployment
Software Engineering
SQL Stored Procedures
SQL Databases
Data Streaming
Systems Integration
Tableau
Workflow Management Systems
YAML
Parquet
Google Cloud Platform
Enterprise Software Applications
Delivery Pipeline
Multi-Agent Systems
Generative AI
Data Lake
PySpark
Information Technology
Avro
REST
Terraform
Looker Analytics
Data Pipelines
Databricks

Job description

We are seeking a skilled Senior Data Engineer with expertise in Google Cloud Platform (Google Cloud Platform), Databricks, SQL, and Python/PySpark. The candidate will be responsible for designing, developing, optimizing, and maintaining scalable data pipelines and modern data platforms. This role requires collaboration with architects, developers, analysts, and business stakeholders while taking ownership of end-to-end data engineering solutions., * Design, develop, and maintain scalable ETL/ELT pipelines on Google Cloud Platform (Google Cloud Platform).

  • Develop complex BigQuery SQL scripts, stored procedures, functions, and Common Table Expressions (CTEs).
  • Build and manage data pipelines using Cloud Composer (Airflow DAGs).
  • Develop streaming and batch data pipelines using Pub/Sub, Dataflow, and Dataproc.
  • Design and implement data processing workflows using Databricks (PySpark and Spark SQL).
  • Develop and maintain GitHub-based CI/CD deployment pipelines.
  • Create and manage Eventarc integrations on Google Cloud Platform.
  • Build data integration pipelines between SQL databases, BigQuery, Databricks, and other enterprise systems.
  • Work with modern data architectures, including Data Lakes and Lakehouse environments.
  • Process and manage data stored in formats such as Parquet, Avro, Apache Iceberg, and Delta Lake.
  • Perform data profiling, exploratory data analysis (EDA), and validation of source and target datasets.
  • Identify data quality issues, anomalies, and inconsistencies, and implement corrective measures.
  • Validate transformation logic against business requirements.
  • Ingest data from relational databases, flat files, REST APIs, and other external sources.
  • Ensure data quality, reconciliation, monitoring, and performance optimization.
  • Troubleshoot and resolve production data pipeline issues.
  • Participate in code reviews and promote data engineering best practices.
  • Collaborate with analytics, reporting, and application development teams.
  • Support production deployments and ongoing enhancements.

Requirements

Education: Bachelor's degree in Computer Science, Information Technology, or a related field (or equivalent experience).

Experience: 8+ years of experience in Data Engineering.

Technical Skills:

  • Expertise in SQL, including complex query writing, SPs, Functions, CTEs, query optimization and performance tuning.
  • Advanced programming skills in Python and PySpark.
  • Hands-on experience with Databricks.
  • Experience with Google Cloud Platform (Google Cloud Platform) services, including Cloud Storage, BigQuery, Cloud Composer, Pub/Sub, Dataflow, Dataproc, Eventarc, and AlloyDB.
  • Experience implementing GitHub-based CI/CD pipelines.
  • Experience with batch and streaming data processing.
  • Understanding of Data Lake and Lakehouse architectures.
  • Experience working with Parquet, Avro, Apache Iceberg, and Delta Lake.
  • Experience performing data profiling, reconciliation, and exploratory data analysis.
  • Familiarity with Generative AI development tools such as OpenAI Codex, Cursor, or similar AI-assisted coding tools.
  • Knowledge of Terraform, including provisioning and managing cloud infrastructure.
  • Knowledge of ETL tools such as Informatica or similar platforms is preferred.

Preferred Qualifications

  • Experience with cloud data migration projects.
  • Experience with workflow orchestration frameworks such as Apache Airflow or Cloud Composer.
  • Familiarity with YAML-based pipeline configuration.
  • Experience implementing CI/CD for data engineering solutions on Github.
  • Experience with designing, developing, or integrating Generative AI agents and multi-agent systems is preferred.
  • Understanding of data governance, security, and data quality frameworks.
  • Experience with BI and visualization tools such as Power BI, Tableau, Looker, or similar platforms.

Soft Skills

  • Analytical and problem-solving skills.
  • Verbal and written communication skills.
  • Ability to work independently as well as collaboratively in cross-functional teams.
  • A proactive attitude with a willingness to learn, take ownership, and drive initiatives to completion.
  • Motivated, with a passion for learning and delivering innovative solutions.

About the company

Everforth Apex is a world-class IT services company that serves thousands of clients across the globe. When you join Everforth Apex, you become part of a team that values innovation, collaboration, and continuous learning. We offer quality career resources, training, certifications, development opportunities, and a comprehensive benefits package. Our commitment to excellence is reflected in many awards, including ClearlyRateds Best of Staffing in Talent Satisfaction in the United States and Great Place to Work in the United Kingdom and Mexico. Everforth Apex uses a virtual recruiter as part of the application process. Click for more details. By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Everforth Apex and its affiliates, and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy at

Apply for this position