Data Engineer- Offshore

Insight Global
Houston, TX, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Compensation
$33,280.0 - $41,600.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Amazon Web Services Big Data Cloud Engineering Computer Programming Information Engineering Data Warehousing Distributed Computing Environment Distributed Systems Apache Hadoop Python (Programming Language) Cloud Services
+13 more
Scala (Programming Language) Software Engineering Data Streaming Enterprise Data Management Data Processing Data Ingestion Apache Spark Software Troubleshooting Pyspark Apache Flume Real Time Data Apache Kafka Video Streaming

Job description

The Senior Data Engineer / Streaming Data Engineer will design, build, and support enterprise-scale streaming and big data pipelines across AWS, Spark, Hadoop, Kafka, and cloud-native ingestion platforms. This role is hands-on and production-focused, with responsibility for reliable real-time data movement, ingestion modernization, distributed systems engineering, and scalable data processing in a large enterprise environment.

  • Hands-on streaming and real-time data engineering using Kafka, Spark, AWS Kinesis, and cloud-native data services.
  • Modernization opportunity focused on moving legacy ingestion patterns to scalable AWS-native services.
  • Production engineering role requiring strong troubleshooting, operational ownership, and distributed systems depth.
  • Enterprise-scale environment with complex data warehousing, big data, and cross-platform integration needs.

Requirements

  • 10+ years of experience in data engineering, software engineering, or related fields.
  • Strong expertise with AWS cloud platform services and cloud-native data engineering patterns.
  • Hands-on experience with Apache Spark using Scala and PySpark for large-scale data processing.
  • Deep working knowledge of the Hadoop ecosystem and distributed data processing architectures.
  • Strong experience with Kafka and streaming technologies, including real-time data pipeline design and support.
  • Hands-on experience with data ingestion platforms such as Flume, AWS Kinesis, Kinesis Firehose, or similar tooling.
  • Strong programming experience in Python, Scala, and Java.
  • Deep understanding of enterprise data warehousing, big data architectures, distributed systems, and large-scale enterprise operating environments.

Benefits & conditions

Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.insightglobal.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

2:04 min

Comparing offline data analytics with online stream processing

Artem Volk Artem Volk +1 · WWC 2024

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

Videos

See all

Related articles

See all