Data Engineer

IBA InfoTech Inc.
Durham, NC, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Java (Programming Language) Computer-Aided Design Amazon Web Services Big Data Cloud Computing Code Review Programming Tools Distributed Systems Elasticsearch Fault Tolerance Apache Hadoop Apache Hive
+13 more
Information Lifecycle Management Python (Programming Language) Data Streaming Apache Spark Information Technology Druid Apache Flink Cassandra Data Analytics Presto Stream Processing Data Pipelines Programming Languages

Job description

  • Develop and deploy highly-available, fault-tolerant software that will help drive improvements towards the features, reliability, performance, and efficiency of the Cloud Analytics platform.
  • Actively review code, mentor, and provide peer feedback.
  • Collaborate with engineering teams to identify and resolve pain points as well as evangelize best practices.
  • Partner with various teams to transform concepts into requirements and requirements into services and tools.
  • Engineer efficient, adaptable and scalable architecture for all stages of data lifecycle (ingest, streaming, structured and unstructured storage, search, aggregation) in support of a variety of data applications.
  • Build abstractions and re-usable developer tooling to allow other engineers to quickly build streaming/batch self-service pipelines.
  • Build, deploy, maintain, and automate large global deployments in AWS.
  • Troubleshoot production issues and come up with solutions as required.

Requirements

  • You have a strong engineering background with ability to design software systems from the ground up.
  • You have expertise in Java, Python or similar programming languages.
  • You have experience in web-scale data and large-scale distributed systems, ideally on cloud infrastructure.
  • You have a product mindset. You are energized by building things that will be heavily used.
  • You have engineered scalable software using big data technologies (e.g. Hadoop, Spark, Hive, Presto, Flink, Samza, Storm, Elasticsearch, Druid, Cassandra, etc).
  • You have experience building data pipelines (real-time or batch) on large complex datasets.
  • You have worked on and understand messaging/queueing/stream processing systems.
  • You design not just with a mind for solving a problem, but also with maintainability, testability, monitorability, and automation as top concerns.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on ibainfotech.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:53 min

Preparing highly scalable facial recognition for big data

Sefik Serengil · LIVE

Videos

See all

Related articles

See all