Hadoop / Big Data Platform Engineer - SME

Ssv Technologies Inc
Charlotte, NC, United States
21 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Query Performance Java (Programming Language) Artificial Intelligence CA Workload Automation Ae Big Data Software Bug Management Cloud Computing Cloudera Impala Encodings Databases Data Infrastructure Extract Transform Load (ETL)
+39 more
Data Warehousing Relational Databases Software Debugging DevOps Disaster Recovery Perl (Programming Language) Apache Hadoop Hadoop Distributed File System Apache HBase Apache Hive Job Scheduling Python (Programming Language) Kerberos (Protocol) PostgreSQL Log Files Node.Js OpenShift Oracle (Applications) Ansible Standard Sql Cloudera Shell Script Simple Data Format Apache Solr SQL Databases Subversion Tableau (Software) Apache Zookeeper Jupyter Notebook Apache Yarn Apache Spark Sybase Code Testing Sentry Apache Kafka Bitbucket Docker Jenkins Databricks

Job description

Hadoop Engineer (SME) role supporting NextGen Platforms built around Big Data Technologies (Hadoop, Spark, Kafka, Impala, HBase, Docker-Container, Ansible and many more). Requires experience in cluster management of vendor-based Hadoop and Data Science (AI/ML) products like Cloudera, Databricks. Hadoop Engineer is involved in the full life cycle of an application and part of an agile development process. They require the ability to interact, develop, engineer, and communicate collaboratively at the highest technical levels with clients, development teams, vendors, and other partners. The following section is intended to serve as a general guideline for each relative dimension of project complexity, responsibility, and education/experience within this role.

  • Works on complex, major or highly visible tasks in support of multiple projects that require multiple areas of expertise
  • Team member will be expected to provide subject matter expertise in managing Hadoop and Data Science Platform operations with focus around Cloudera Hadoop, Jupyter Notebook, OpenShift, Docker-Container Cluster Management and Administration
  • Integrates solutions with other applications and platforms outside the framework
  • He / She will be responsible for managing platform operations across all environments which includes upgrades, bug fixes, deployments, metrics / monitoring for resolution and forecasting, disaster recovery, incident / problem / capacity management
  • Serves as a liaison between client partners and vendors in coordination with project managers to provide technical solutions that address user needs
  • Experience on L1, L2 L3 level support in Hadoop Platform
  • Experience in Yarn , Spark and Impala job debugging and troubleshooting
  • Experience in addressing issues around Name node , HDFS Space , File/Folder Permission backups

Requirements

  • Understanding and experience in experience in query performance tuning and resource utilization tuning
  • Technical knowledge on relational databases, data warehousing and SQL skills
  • Hadoop, Kafka, Spark, Impala, Hive, HBase, Ozone etc.
  • Strong knowledge of Hadoop Architecture, HDFS, Hadoop Cluster and Hadoop Administrator’’s role
  • Intimate knowledge of fully integrated AD/Kerberos authentication
  • Experience setting up optimum cluster configurations
  • Debugging knowledge of YARN.
  • Hands-on with analyzing various Hadoop log files, compression, encoding, file formats
  • Expert level knowledge of Cloudera Hadoop components such as HDFS, Sentry, HBase, Kafka, Impala, SOLR, Hue, Spark, Hive, YARN, Zookeeper and Postgres
  • Strong technical knowledge: Unix/Linux; Database (Sybase/SQL/Oracle), Java, Python, Perl, Shell scripting, Infrastructure.
  • Experience in Monitoring Alerting, and Job Scheduling Systems
  • Being comfortable with frequent, incremental code testing and deployment
  • Strong grasp of automation / DevOps tools Ansible, Jenkins, SVN, Bitbucket

Tools Involved - ETL tools: Hadoop Stack Job Scheduling tools: Autosys BI tools: Tableau ; Tableau dashboard building connecting Hadoop-Hive Tables is equally preferred Shift Timings: 8AM 5PM / 12:30PM ET 9:30PM ET Desired Qualifications: Experience working on Big Data Technologies Cloudera Admin / Dev Certification Certification in Cloud, Docker-Container, OpenShift Technologies

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:22 min

Overview of the Sentry error and performance monitoring platform

Priscila Oliveira · World Congress 2023

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:21 min

Installing and configuring the Sentry JavaScript SDK for applications

Priscila Oliveira · World Congress 2023

Videos

See all

Related articles

See all