Hadoop Big Data Developer

Bright Vision Technologies
Burlington, United States of America
yesterday

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
$ 150K

Job location

Remote
Burlington, United States of America

Tech stack

Java
Airflow
Apache HTTP Server
Azure
Big Data
Cloud Computing
Continuous Integration
Information Engineering
Data Governance
Data Stores
Database Queries
Software Debugging
Distributed Systems
Fault Tolerance
Hadoop
Hadoop Distributed File System
HBase
Hive
Python
Machine Learning
NoSQL
Apache Oozie
Sqoop
Data Streaming
Unstructured Data
Scripting (Bash/Python/Go/Ruby)
Spark
Hdinsight
Kubernetes
Information Technology
Collibra
Apache Flink
Kafka
Spark Streaming
Data Management
Amazon Web Services (AWS)
Databricks

Job description

We are seeking an experienced Hadoop Big Data Developer to design, build, and operate large-scale data processing pipelines and analytics platforms on Hadoop and related big-data ecosystems. In this role you will be responsible for ingesting, transforming, and analyzing massive volumes of structured and unstructured data to support enterprise analytics, machine learning, and reporting workloads. The ideal candidate will combine deep technical expertise across the Hadoop ecosystem with strong software engineering fundamentals and a clear understanding of how to deliver reliable, performant, and cost-effective data platforms in production environments.

Requirements

  • Bachelor's degree in Computer Science, Engineering, or a related technical discipline.
  • Five or more years of professional experience designing and operating big-data pipelines on Hadoop.
  • Strong hands-on expertise with Apache Spark (Scala, Python, or Java) in production environments.
  • Solid experience with Hive, HDFS, Sqoop, HBase, and the broader Hadoop ecosystem.
  • Hands-on experience with streaming data platforms such as Kafka, Spark Streaming, or Flink.
  • Strong SQL skills and experience working with both relational and NoSQL data stores.
  • Experience with workflow orchestration tools such as Airflow or Oozie.
  • Solid understanding of distributed systems concepts, including partitioning, replication, and fault tolerance.
  • Strong scripting skills in Python or Shell.
  • Excellent troubleshooting, debugging, and documentation skills.

Preferred Qualifications

  • Experience operating Hadoop on cloud platforms such as AWS EMR, Azure HDInsight, or Databricks.
  • Familiarity with modern lakehouse formats (Delta, Iceberg, Hudi).
  • Exposure to data governance tooling such as Apache Atlas or Collibra.
  • Experience with Kubernetes-based data platforms (Spark-on-K8s, Trino).
  • Hands-on experience with CI/CD and infrastructure-as-code in data engineering workflows.

About the company

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

Apply for this position