> Markdown version of [/jobs/ext/2044136-data-engineer](https://www.wearedevelopers.com/jobs/ext/2044136-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Peoplentech Llc - **Location:** Santa Clara, CA, United States - **Salary:** $145,600.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Agile Methodology, Airflow, Amazon Web Services, Data Analysis, Google App Engines, Automation of Tests, Unit Testing, Microsoft Azure, BigQuery, Cloud Computing, Cloud Database, Configuration Management, Databases, Continuous Delivery, Continuous Integration, Information Engineering, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Migration, Data Systems, Data Warehousing, DevOps, Disaster Recovery, Amazon DynamoDB, Data Flow Control, Apache Hive, Python (Programming Language), NoSQL, Performance Tuning, Regression Testing, Migration Manager, Requirements Management, Scala (Programming Language), Software Deployment, Software Engineering, Data Streaming, Systems Integration, Virtual Machines, Private Cloud Environment, Data Processing, Scripting, Cloud Platform System, System Availability, Delivery Pipeline, Snowflake, Apache Spark, Electronic Medical Records, AWS Lambda, Amazon Virtual Private Cloud (VPC), Git, Cloudformation, Gitlab-ci, Integration Tests, Kubernetes, Low Latency, AWS Glue, Data Analytics, Star Schema, AWS Data Analytics, Real Time Data, Apache Kafka, Data Management, Multiplatform, Data Pipelines, Serverless Computing, Amazon Elastic Mapreduce (EMR), Docker, Jenkins, Amazon Redshift, Databricks - **Published:** August 13, 2026 - **Apply:** https://www.careerbuilder.com/job-details/data-engineer-santa-clara-ca--df7ccd81-0ec0-4c99-9784-d734048be983 ## About the Role Languages & Scripting: Spark, Python, Java, Scala, Hive, Kafka, SQL Cloud Platforms: AWS Data Warehousing & Analytics: Redshift or Snowflake or Big Query Data Integration & ETL: AWS Glue, Aws EMR, Spark, Data Bricks CI/CD: AWS Code Pipeline, Jenkins, CloudFormation, Docker, Kubernetes JD : * Results-driven Data Engineer with a decade of expertise in Data engineering across cloud platforms with a total of 12 years in IT. * Extensive experience utilizing Google Cloud Platform (GCP) services, including BigQuery, Dataflow, Dataprep, and Pub/Sub, for data engineering solutions. * Proficient in building and managing GCP data pipelines with tools like Cloud Composer and Cloud Dataflow. * Proven ability in developing and deploying applications on Google Kubernetes Engine (GKE). * Strong background in implementing security and compliance on GCP, ensuring data privacy and regulatory adherence. * Track record of optimizing cost and resource usage within GCP environments. * Skilled in AWS services such as Amazon EMR, Redshift, and Glue for efficient data processing. * Expertise in architecting scalable, cost-effective solutions on AWS, with proficiency in configuring AWS Lambda for serverless computing. * Adept at setting up AWS Kinesis streams to process real-time data, enhancing system responsiveness and data-driven decision-making. * Proficient in leveraging AWS DynamoDB to create scalable, low-latency NoSQL databases for dynamic applications. * Deep expertise in optimizing and managing Amazon Redshift data warehouses to deliver high-performance analytics and business insights. * Experienced in integrating AWS services into CI/CD pipelines, streamlining automation for continuous integration, delivery, and deployment. * Skilled in setting up and securing AWS Virtual Private Cloud (VPC) environments. * Proficient in managing Azure virtual machines (VMs) for cloud infrastructure operations. * Extensive experience managing on-premises data infrastructure, including data warehouses and databases. * Familiar with AWS DevOps practices for continuous integration and deployment. * Expertise in using Git for version control in DBT projects, ensuring proper tracking and documentation of data model changes. * Skilled in performance optimization and tuning of on-premises data systems. * Proficient in data migration strategies between on-premises and cloud environments. * Strong troubleshooting skills in resolving issues within on-premises data systems. * Proven ability to maintain high availability and disaster recovery solutions in on-premises environments. * Experienced in implementing CI/CD pipelines using tools like Jenkins and GitLab CI/CD. * Adept in automated testing processes, including unit, integration, and regression testing. * Skilled in gathering and analyzing project requirements to ensure alignment with business goals. * Experienced in Agile project management, contributing to successful outcomes through data-driven analytics and collaborative teamwork Skills: AWS Lambda, Agile Programming Methodologies, Amazon Web Services (AWS), Automation, Cloud Computing, Continuous Deployment/Delivery, Continuous Integration, Cost Control, Data Analysis, Data Management, Data Migration, Data Modeling, Data Processing, Data Warehousing, Database Extract Transform and Load (ETL), DevOps, Disaster Recovery, Docker, Documentation Models, Electronic Medical Records, GCP (Good Clinical Practices), Git, Google App Engine (GAE), High Availability, Identify Issues, Integration Testing, Jenkins, Microsoft Windows Azure, Migration Strategy, Multiplatform/Cross-Platform, NoSQL, Performance Analysis, Performance Tuning/Optimization, Privacy Regulations, Private Cloud, Problem Solving Skills, Project Control, Project Evaluation, Project/Program Management, Quality Assurance Methodology, Regression Testing, Requirements Management, Sales Pipeline, Snowflake Schema, Software Development, Software Engineering, Source Code/Configuration Management (SCM), Team Player, Test Automation, Testing, Unit Test, VMS Operating System, Virtual Machine (VM) ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers)