Data Architect

Apptad Inc.
Plano, TX, United States
about 1 month ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$124,200.0 - $175,800.0
Working hours
Regular working hours

Tech stack

Airflow Amazon S3 Apache HTTP Server Batch Processing Big Data Cloud Computing Cloudera Impala Continuous Integration Data Architecture Data Governance Data Security Software Design Patterns
+30 more
Disaster Recovery Memory Management Apache Hadoop Hadoop Distributed File System MapReduce Monitoring of Systems Apache Hive Kerberos (Protocol) Meta-Data Management Performance Tuning Cloudera SQL Databases Data Streaming Backup and Restore Workflow Management Systems Parquet Data Logging Apache Yarn System Availability Apache Spark Software Troubleshooting Change Data Capture Data Lakes Data Lineage Optimization Algorithms AWS Glue Apache Kafka Data Pipelines Amazon Elastic Mapreduce (EMR) Databricks

Requirements

Seeking a Senior Specialist with 7 to 11 years of experience in Data Architecture to provide datadriven solutions Job Description Looking for someone who is familiar with the below areas and Design develop and maintain robust data architecture frameworks to support data requirements and translate them into scalable data solutions HDFS Architecture NameNode High Availability HA HDFS Federation Rack Awareness HDFS ReadWrite Flow YARN Architecture Capacity Scheduler Fair Scheduler Resource Management Queue Design MapReduce Architecture Shuffle and Sort Apache Spark Architecture Spark DAG Execution Plan Spark Performance Tuning Memory Management in Spark Adaptive Query Execution AQE Hive Architecture Hive Metastore Hive Optimization Techniques Partitioning Bucketing ORC Parquet Internals Impala Architecture Apache Kafka Architecture Kafka Performance Tuning Kafka Security Streaming Architectures CDC Change Data Capture Apache Airflow Architecture Workflow Orchestration Data Lake Architecture Lakehouse Architecture Delta Lake Fundamentals Data Modeling for Big Data ETLELT Framework Design Data Governance Data Lineage Metadata Management Kerberos Authentication Apache Ranger Apache Knox Encryption Security Best Practices Cloudera CDP Architecture AWS EMR AWS Glue S3 Architecture Optimization Databricks Architecture CICD for Data Pipelines Monitoring Observability Logging Frameworks Cluster Sizing Capacity Planning Disaster Recovery DR Backup Recovery Strategies High Availability Design MultiTenancy Design Data Quality Frameworks Performance Troubleshooting Data Skew Handling Join Optimization Techniques Small File Problem Cost Optimization RealTime Data Pipeline Design Batch Processing Architecture EventDriven Architecture Lambda Architecture EndtoEnd Solution Architecture Migration from OnPrem Hadoop to Cloud Scalability Reliability Design Patterns System Design for Big Data Platforms Architecture Tradeoffs and Decision Making SQL PySparkScala Coding

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

55 sec

Validating data processing architectures via containerized events

Modood Alvi · World Congress 2025

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:03 min

Introduction to open table formats built on Parquet

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

1:09 min

Evaluating mature stream processing frameworks for production systems

Soroosh Khodami Soroosh Khodami · World Congress 2024

Videos

See all

Related articles

See all