Hadoop Engineer

Randstad
Plano, TX, United States
22 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$124,280.0 - $145,080.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Artificial Intelligence Big Data Software Bug Management Cloud Computing Cloudera Impala Encodings Databases Software Debugging DevOps Disaster Recovery Perl (Programming Language)
+32 more
Apache Hadoop Hadoop Distributed File System Apache HBase Apache Hive Job Scheduling Python (Programming Language) Kerberos (Protocol) PostgreSQL Log Files OpenShift Oracle (Applications) Ansible Cloudera Shell Script Simple Data Format Apache Solr SQL Databases Subversion Talend Apache Zookeeper Jupyter Notebook Apache Yarn Snowflake Apache Spark Sybase Code Testing Sentry Apache Kafka Bitbucket Docker Jenkins Databricks

Job description

Hadoop Engineer (SME) role supporting NextGen Platforms built around Big Data Technologies (Hadoop, Spark, Kafka, Impala, Hbase, Docker-Container, Ansible and many more). Requires experience in cluster management of vendor based Hadoop and Data Science (AI/ML) products like Cloudera, Databricks, Snowflake, Talend, Greenfield, ELK, KPMG Ignite etc. Hadoop Engineer is involved in the full life cycle of an application and part of an agile development process. They require the ability to interact, develop, engineer, and communicate collaboratively at the highest technical levels with clients, development teams, vendors and other partners. The following section is intended to serve as a general guideline for each relative dimension of project complexity, responsibility and education/experience within this role., Hadoop Engineer (SME) role supporting NextGen Platforms built around Big Data Technologies (Hadoop, Spark, Kafka, Impala, Hbase, Docker-Container, Ansible and many more). Requires experience in cluster management of vendor based Hadoop and Data Science (AI/ML) products like Cloudera, Databricks, Snowflake, Talend, Greenfield, ELK, KPMG Ignite etc. Hadoop Engineer is involved in the full life cycle of an application and part of an agile development process. They require the ability to interact, develop, engineer, and communicate collaboratively at the highest technical levels with clients, development teams, vendors and other partners. The following section is intended to serve as a general guideline for each relative dimension of project complexity, responsibility and education/experience within this role.

qualifications:

Works on complex, major or highly visible tasks in support of multiple projects that require multiple areas of expertise

Team member will be expected to provide subject matter expertise in managing Hadoop and Data Science Platform operations with focus around Cloudera Hadoop, Jupyter Notebook, OpenShift, Docker-Container Cluster Management and Administration

Integrates solutions with other applications and platforms outside the framework

He / She will be responsible for managing platform operations across all environments which includes upgrades, bug fixes, deployments, metrics / monitoring for resolution and forecasting, disaster recovery, incident / problem / capacity management

Serves as a liaison between client partners and vendors in coordination with project managers to provide technical solutions that address user needs

Hadoop, Kafka, Spark, Impala, Hive, Hbase etc.

Requirements

Strong knowledge of Hadoop Architecture, HDFS, Hadoop Cluster and Hadoop Administrator’s role

Intimate knowledge of fully integrated AD/Kerberos authentication

Experience setting up optimum cluster configurations

Debugging knowledge of YARN.

Hands-on with analyzing various Hadoop log files, compression, encoding, file formats

Expert level knowledge of Cloudera Hadoop components such as HDFS, Sentry, HBase, Kafka, Impala, SOLR, Hue, Spark, Hive, YARN, Zookeeper and Postgres

Strong technical knowledge: Unix/Linux; Database (Sybase/SQL/Oracle), Java, Python, Perl, Shell scripting, Infrastructure.

Experience in Monitoring Alerting, and Job Scheduling Systems

Being comfortable with frequent, incremental code testing and deployment

Strong grasp of automation / DevOps tools - Ansible, Jenkins, SVN, Bitbucket

Experience working on Big Data Technologies

Cloudera Admin / Dev Certification

Certification in Cloud, Docker-Container, OpenShift Technologies

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

1:22 min

Overview of the Sentry error and performance monitoring platform

Priscila Oliveira · WWC 2023

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

3:21 min

Installing and configuring the Sentry JavaScript SDK for applications

Priscila Oliveira · WWC 2023

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all