Data Engineer

Robert Half
The Woodlands, TX, United States
3 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Big Data Extract Transform Load (ETL) Data Transformation Distributed Computing Environment Distributed Systems Apache Hadoop Python (Programming Language) DataOps Data Streaming Data Processing System Availability Apache Spark
+2 more
Apache Kafka Data Pipelines

Job description

Description We are looking for a Data Engineer to support scalable data solutions for a long-term contract opportunity in The Woodlands, Texas. This role focuses on designing and optimizing data pipelines, integrating large-scale data sources, and enabling reliable access to critical business information. The ideal candidate will bring strong hands-on experience with modern big data technologies and a practical approach to building efficient ETL workflows.

Responsibilities:

  • Build, maintain, and enhance robust data pipelines to process large volumes of structured and unstructured information.

  • Develop ETL workflows that transform raw data into reliable datasets for analytics, reporting, and operational use.

  • Use Python and Apache Spark to engineer high-performance data processing solutions across distributed environments.

  • Work with Apache Hadoop ecosystems to manage storage and support scalable data operations.

  • Integrate streaming and event-driven data using Apache Kafka to improve data availability and timeliness.

  • Monitor data workflows, troubleshoot processing issues, and implement improvements that increase reliability and efficiency.

  • Collaborate with technical and business stakeholders to understand data needs and translate them into practical engineering solutions.

Requirements

  • Document pipeline architecture, data flow logic, and operational procedures to support maintainability and team knowledge sharing. Requirements * Proven experience working as a Data Engineer in environments that handle large and complex datasets.

  • Strong hands-on expertise with Python for data engineering, automation, and pipeline development.

  • Practical experience using Apache Spark for distributed data processing and transformation.

  • Familiarity with Apache Hadoop and related big data frameworks.

  • Experience with Apache Kafka for data streaming or message-based integration.

  • Solid understanding of ETL design, data transformation practices, and data quality principles.

  • Ability to diagnose technical issues, optimize data workflows, and deliver dependable solutions in a collaborative setting. Technology Doesn’t Change the World, People Do.®

About the company

Robert Half is the world’s first and largest specialized talent solutions firm that connects highly qualified job seekers to opportunities at great companies. We offer contract, temporary and permanent placement solutions for finance and accounting, technology, marketing and creative, legal, and administrative and customer support roles.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:00 min

Separating dataset creation from low-level software implementation steps

Jan Zawadzki · World Congress 2022

1:19 min

The true role and evolution of data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

Videos

See all

Related articles

See all