> Markdown version of [/jobs/ext/2705558-hadoop-architect-data-engineer-capacity-analytics-architect](https://www.wearedevelopers.com/jobs/ext/2705558-hadoop-architect-data-engineer-capacity-analytics-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Hadoop Architect + Data Engineer + Capacity Analytics Architect - **Company:** BCforward - **Location:** United States - **Experience:** Expert - **Salary:** $153,234.0 - **Contract:** Temporary contract - **Skills:** Adaptable Database Systems, Application Programming Interfaces (APIs), Amazon S3, Business Analytics Applications, CA Workload Automation Ae, Cloud Storage, Cloudera Impala, Data as a Services, Extract Transform Load (ETL), Data Systems, Data Warehousing, Relational Databases, Database Queries, Apache Hadoop, Hadoop Distributed File System, Apache HBase, Apache Hive, Python (Programming Language), CURL, PostgreSQL, Windows PowerShell, Power BI, Prometheus, Cloudera, Shell Script, Tableau (Software), Teradata SQL, Enterprise Data Management, Scripting, Apache Yarn, Cloudera Manager, Grafana, Advanced Reports, Powerquery, Pyspark, Vba Programming Language, Data Analytics, Api Design - **Published:** September 4, 2026 - **Apply:** https://www.dice.com/job-detail/2216475b-51cb-4679-b1a5-4781ffbf946d ## About the Role We are seeking a Senior Data Solutions Architect to join our dynamic Enterprise Data Platforms Advanced Analytics team. The ideal candidate will have strong experience in Hadoop architecture, enterprise data lakes, capacity and resource analytics, and API-driven ETL and a proven ability to design and deliver scalable analytics solutions that drive platform insights, automate reporting, and enable data-driven decisions., * Deep Hadoop architecture expertise, including HDFS, HBase, Ozone, MinIO, Hive, Impala, YARN, and Hadoop CLI. * Proficiency with FSImage analysis and storage compute utilization assessment. * Hands-on API programming and ETL using cURL, PySpark, Python, shell scripting, and PowerShell. * Experience with Autosys for scheduling, monitoring, and error resolution. * Strong data modeling, advanced query writing, and analytical dataset development skills. * Experience with Cloudera data services, Cloudera Manager, Pepperdata, and/or Acceldata Pulse. * Familiarity with object storage and S3-compatible platforms, bucket-based storage models, and cloud storage concepts. * Proficiency in Excel automation using VBA, Power Query, and scripting; familiarity with Tableau and Power BI. * RDBMS experience, including Teradata and PostgreSQL. * Demonstrated ability to operate independently with ownership, attention to detail, and a continuous improvement mindset. Preferred Skills: * Experience with Prometheus and Grafana for advanced platform visibility and trend analysis. * Experience in cloud or hybrid data ecosystems and compute analytics. ## Description * Build analytics solutions within the Enterprise Data Warehouse and enterprise data lake to deliver capacity and resource insights. * Develop a single source of truth for capacity, storage, compute, onboarding demand, and platform hygiene data. * Design scalable data models and analytical datasets across Hadoop and cloud platforms. * Automate reporting and dashboards to support leadership decision-making. * Analyze CPU, memory, and workload utilization using YARN, Hive, and Impala metrics for resource reporting. * Leverage HDFS CLI and FSImage analysis to assess storage and compute utilization, including Ozone and object storage platforms. * Develop API-driven ETL pipelines using cURL, PySpark, and shell scripting with Autosys scheduling and monitoring. * Use Cloudera Manager, Pepperdata, and/or Acceldata Pulse for platform monitoring and operational insights. * Translate complex data into actionable insights and executive-ready narratives. * Lead POCs for emerging technologies and recommend scalable implementation strategies. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Don’t Insert Crazy! On cURL and AI Slop - Daniel Stenberg](https://www.wearedevelopers.com/videos/1796-don-t-insert-crazy-on-curl-and-ai-slop-daniel-stenberg) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)