> Markdown version of [/jobs/ext/1509789-data-architect](https://www.wearedevelopers.com/jobs/ext/1509789-data-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Architect - **Company:** Apptad Inc. - **Location:** Plano, TX, United States - **Experience:** Expert - **Salary:** $124,200.0 - $175,800.0 - **Contract:** Permanent contract - **Skills:** Airflow, Amazon S3, Apache HTTP Server, Batch Processing, Big Data, Cloud Computing, Cloudera Impala, Continuous Integration, Data Architecture, Data Governance, Data Security, Software Design Patterns, Disaster Recovery, Memory Management, Apache Hadoop, Hadoop Distributed File System, MapReduce, Monitoring of Systems, Apache Hive, Kerberos (Protocol), Meta-Data Management, Performance Tuning, Cloudera, SQL Databases, Data Streaming, Backup and Restore, Workflow Management Systems, Parquet, Data Logging, Apache Yarn, System Availability, Apache Spark, Software Troubleshooting, Change Data Capture, Data Lakes, Data Lineage, Optimization Algorithms, AWS Glue, Apache Kafka, Data Pipelines, Amazon Elastic Mapreduce (EMR), Databricks - **Published:** July 30, 2026 - **Apply:** https://www.careerjet.com/jobad/uscc0185baf176e0efd10d02e69042501a ## About the Role Seeking a Senior Specialist with 7 to 11 years of experience in Data Architecture to provide datadriven solutions Job Description Looking for someone who is familiar with the below areas and Design develop and maintain robust data architecture frameworks to support data requirements and translate them into scalable data solutions HDFS Architecture NameNode High Availability HA HDFS Federation Rack Awareness HDFS ReadWrite Flow YARN Architecture Capacity Scheduler Fair Scheduler Resource Management Queue Design MapReduce Architecture Shuffle and Sort Apache Spark Architecture Spark DAG Execution Plan Spark Performance Tuning Memory Management in Spark Adaptive Query Execution AQE Hive Architecture Hive Metastore Hive Optimization Techniques Partitioning Bucketing ORC Parquet Internals Impala Architecture Apache Kafka Architecture Kafka Performance Tuning Kafka Security Streaming Architectures CDC Change Data Capture Apache Airflow Architecture Workflow Orchestration Data Lake Architecture Lakehouse Architecture Delta Lake Fundamentals Data Modeling for Big Data ETLELT Framework Design Data Governance Data Lineage Metadata Management Kerberos Authentication Apache Ranger Apache Knox Encryption Security Best Practices Cloudera CDP Architecture AWS EMR AWS Glue S3 Architecture Optimization Databricks Architecture CICD for Data Pipelines Monitoring Observability Logging Frameworks Cluster Sizing Capacity Planning Disaster Recovery DR Backup Recovery Strategies High Availability Design MultiTenancy Design Data Quality Frameworks Performance Troubleshooting Data Skew Handling Join Optimization Techniques Small File Problem Cost Optimization RealTime Data Pipeline Design Batch Processing Architecture EventDriven Architecture Lambda Architecture EndtoEnd Solution Architecture Migration from OnPrem Hadoop to Cloud Scalability Reliability Design Patterns System Design for Big Data Platforms Architecture Tradeoffs and Decision Making SQL PySparkScala Coding ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [How we built an AI-powered code reviewer in 80 hours](https://www.wearedevelopers.com/videos/1511-how-we-built-an-ai-powered-code-reviewer-in-80-hours) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Tips, Techniques, and Common Pitfalls Debugging Kafka](https://www.wearedevelopers.com/videos/838-tips-techniques-and-common-pitfalls-debugging-kafka) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Why Event-Driven Architecture Isn’t About Speed (and When You Actually Need It)](https://www.wearedevelopers.com/magazine/745-why-event-driven-architecture-isn-t-about-speed-and-when-you-actually-need-it)