Data Engineer

Peoplentech Llc
Santa Clara, CA, United States
4 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$145,600.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Agile Methodology Airflow Amazon Web Services Data Analysis Google App Engines Automation of Tests Unit Testing Microsoft Azure BigQuery Cloud Computing Cloud Database
+60 more
Configuration Management Databases Continuous Delivery Continuous Integration Information Engineering Data Infrastructure Data Integration Extract Transform Load (ETL) Data Migration Data Systems Data Warehousing DevOps Disaster Recovery Amazon DynamoDB Data Flow Control Apache Hive Python (Programming Language) NoSQL Performance Tuning Regression Testing Migration Manager Requirements Management Scala (Programming Language) Software Deployment Software Engineering Data Streaming Systems Integration Virtual Machines Private Cloud Environment Data Processing Scripting Cloud Platform System System Availability Delivery Pipeline Snowflake Apache Spark Electronic Medical Records AWS Lambda Amazon Virtual Private Cloud (VPC) Git Cloudformation Gitlab-ci Integration Tests Kubernetes Low Latency AWS Glue Data Analytics Star Schema AWS Data Analytics Real Time Data Apache Kafka Data Management Multiplatform Data Pipelines Serverless Computing Amazon Elastic Mapreduce (EMR) Docker Jenkins Amazon Redshift Databricks

Requirements

Languages & Scripting: Spark, Python, Java, Scala, Hive, Kafka, SQL Cloud Platforms: AWS Data Warehousing & Analytics: Redshift or Snowflake or Big Query Data Integration & ETL: AWS Glue, Aws EMR, Spark, Data Bricks CI/CD: AWS Code Pipeline, Jenkins, CloudFormation, Docker, Kubernetes

JD :

  • Results-driven Data Engineer with a decade of expertise in Data engineering across cloud platforms with a total of 12 years in IT.
  • Extensive experience utilizing Google Cloud Platform (GCP) services, including BigQuery, Dataflow, Dataprep, and Pub/Sub, for data engineering solutions.
  • Proficient in building and managing GCP data pipelines with tools like Cloud Composer and Cloud Dataflow.
  • Proven ability in developing and deploying applications on Google Kubernetes Engine (GKE).
  • Strong background in implementing security and compliance on GCP, ensuring data privacy and regulatory adherence.
  • Track record of optimizing cost and resource usage within GCP environments.
  • Skilled in AWS services such as Amazon EMR, Redshift, and Glue for efficient data processing.
  • Expertise in architecting scalable, cost-effective solutions on AWS, with proficiency in configuring AWS Lambda for serverless computing.
  • Adept at setting up AWS Kinesis streams to process real-time data, enhancing system responsiveness and data-driven decision-making.
  • Proficient in leveraging AWS DynamoDB to create scalable, low-latency NoSQL databases for dynamic applications.
  • Deep expertise in optimizing and managing Amazon Redshift data warehouses to deliver high-performance analytics and business insights.
  • Experienced in integrating AWS services into CI/CD pipelines, streamlining automation for continuous integration, delivery, and deployment.
  • Skilled in setting up and securing AWS Virtual Private Cloud (VPC) environments.
  • Proficient in managing Azure virtual machines (VMs) for cloud infrastructure operations.
  • Extensive experience managing on-premises data infrastructure, including data warehouses and databases.
  • Familiar with AWS DevOps practices for continuous integration and deployment.
  • Expertise in using Git for version control in DBT projects, ensuring proper tracking and documentation of data model changes.
  • Skilled in performance optimization and tuning of on-premises data systems.
  • Proficient in data migration strategies between on-premises and cloud environments.
  • Strong troubleshooting skills in resolving issues within on-premises data systems.
  • Proven ability to maintain high availability and disaster recovery solutions in on-premises environments.
  • Experienced in implementing CI/CD pipelines using tools like Jenkins and GitLab CI/CD.
  • Adept in automated testing processes, including unit, integration, and regression testing.
  • Skilled in gathering and analyzing project requirements to ensure alignment with business goals.
  • Experienced in Agile project management, contributing to successful outcomes through data-driven analytics and collaborative teamwork

Skills: AWS Lambda, Agile Programming Methodologies, Amazon Web Services (AWS), Automation, Cloud Computing, Continuous Deployment/Delivery, Continuous Integration, Cost Control, Data Analysis, Data Management, Data Migration, Data Modeling, Data Processing, Data Warehousing, Database Extract Transform and Load (ETL), DevOps, Disaster Recovery, Docker, Documentation Models, Electronic Medical Records, GCP (Good Clinical Practices), Git, Google App Engine (GAE), High Availability, Identify Issues, Integration Testing, Jenkins, Microsoft Windows Azure, Migration Strategy, Multiplatform/Cross-Platform, NoSQL, Performance Analysis, Performance Tuning/Optimization, Privacy Regulations, Private Cloud, Problem Solving Skills, Project Control, Project Evaluation, Project/Program Management, Quality Assurance Methodology, Regression Testing, Requirements Management, Sales Pipeline, Snowflake Schema, Software Development, Software Engineering, Source Code/Configuration Management (SCM), Team Player, Test Automation, Testing, Unit Test, VMS Operating System, Virtual Machine (VM)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all