Technology Lead | Big Data - Hadoop | Hadoop

Ipolarity LLC
Des Peres, MO, United States
1 day ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Airflow Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Data Analysis Audit Trail Bash Shell Big Data Cloud Computing Security Continuous Integration Data Infrastructure
+38 more
Linux DevOps Github Apache Hadoop Monitoring of Systems Identity and Access Management Python (Programming Language) Performance Tuning Red Hat Enterprise Linux Ansible Prometheus Cloudera Shell Script Data Streaming Transport Layer Security Google Cloud Cloud Platform System System Availability Grafana Apache Spark Infrastructure as Code (IaC) Amazon Virtual Private Cloud (VPC) Git Cloudformation Containerization Gitlab-ci Kubernetes Infrastructure Automation Frameworks Information Technology AWS Glue Apache Kafka CruiseControl Cloudwatch Terraform Splunk Software Version Control Docker Jenkins

Job description

Experienced Big Data Administrator responsible for managing, supporting, automating, and optimizing enterprise-scale Hadoop, Kafka, AWS cloud, and DevOps platforms. The role focuses on ensuring high availability, performance, security, scalability, and operational excellence across big data ecosystems while collaborating with development, architecture, and business teams., Kafka Administration Deploy, configure, and manage Apache Kafka clusters and AWS MSK environments. Monitor broker health, partitions, replication factors, and consumer lag. Perform capacity planning and cluster scaling activities. Manage Kafka security using SSL, SASL, ACLs, and encryption standards. Troubleshoot producer, consumer, and broker performance issues. Support Kafka Connect, Schema Registry, Cruise Control, and MirrorMaker implementations. AWS Cloud Administration Manage cloud infrastructure services including EC2, S3, IAM, VPC, EBS, CloudWatch, CloudTrail and AWS Glue. Support AWS Managed Streaming for Kafka (MSK), EMR, Lambda, and Airflow environments. Implement cloud security best practices and governance controls. Perform infrastructure provisioning and automation using Infrastructure as Code (IaC). Monitor cloud resource utilization and optimize operational costs. DevOps & Automation Design and maintain CI/CD pipelines using Jenkins, GitHub Actions, GitLab CI, or similar tools. Automate infrastructure deployment using Terraform, CloudFormation, and Ansible. Manage source control repositories and release processes. Implement monitoring and alerting solutions using Prometheus, Grafana, Splunk, ELK, or CloudWatch. Support containerization technologies such as Docker and Kubernetes. Develop automation scripts using Python, Shell, or Bash. Operations & Support Provide Level 2 and Level 3 production support. Participate in on-call support rotations and incident management activities. Perform root cause analysis (RCA) and implement preventive measures. Create and maintain operational documentation and standard operating procedures. Ensure compliance with security, audit, and regulatory requirements.

Requirements

Apache Kafka Administration KRAFT and MSK AWS Cloud Services Linux (RHEL/Rocky Linux) Shell Scripting and Python Jenkins, Git, Ansible, Terraform Docker and Kubernetes Monitoring Tools (Grafana, Prometheus, Splunk) Networking, Security, and High Availability Concepts Performance Tuning and Capacity Planning Preferred Qualifications: Bachelor s degree in Computer Science, Information Technology, or related field. Experience with Cloudera CDP, AWS MSK, Airflow, and Spark. AWS, GCP, Kafka, or Kubernetes certifications. Experience supporting large-scale production environments handling petabyte-scale data workloads. Key Achievements Expected Maintain platform availability above 99.9%. Adopt AI-assisted engineering practices to improve operational efficiency, reduce manual effort, and accelerate troubleshooting and documentation. Automate repetitive operational tasks. Improve cluster performance and resource utilization. Ensure secure, scalable, and reliable data platform operations. Support enterprise data engineering, analytics, and AI/ML workloads efficiently. Minimum years of experience 8-10 years Certifications Needed :No Top 3 responsibilities you would expect the Subcon to shoulder and execute Perform infrastructure provisioning and automation using Infrastructure as Code IaC. Design and maintain CICD pipelines using Jenkins, GitHub Actions, GitLab CI, or similar tools. Automate infrastructure deployment using Terraform, CloudFormation, and Ansible Manage source control repositories and release processes. Implement monitoring and alerting solutions using Prometheus, Grafana, Splunk, ELK, or CloudWatch

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all